It is a DDPG tutorial, we need about 300s for training. I hate OU Process because of its lots of hyper-parameters. So this DDPG has no OU Process. This simplify DDPG can't work well on harder task. Other RL algorithms can work well on harder task but complicated. You can change this
()
source not stored for this graph (policy: none)
nothing calls this directly
no test coverage detected