Applying DDPG Algorithm to Swing-Up and Balance Control for a Double Inverted Pendulum on a Cart
T.T. Ho, Thanh-Sang Tat, Hoang-Anh Ngo, Truong-Son Nguyen et autres
In this study, we apply the Deep Deterministic Policy Gradient (DDPG) algorithm in reinforcement learning to control a double inverted pendulum on a cart (DIPC)- a high order single input-multi output (SIMO) system . The simulation results demonstrate DDPG's stability and effectiveness …
vn (code pays fourni par la source)