Double Deep Q-Network for Trajectory Generation of a Commercial 7DOF Redundant Manipulator
Enrico Marchesini, Davide Corsi, Andrea Benfatti, Alessandro Farinelli, Paolo Fiorini
- 发表年份
- 2019
- 引用次数
- 6
摘要
Recent studies to solve industrial automation applications where a new and unfamiliar environment is presented have seen a shift to solutions using Reinforcement Learning policies. Building upon the recent success of Deep Q-Networks (DQNs), we present a comparison between DQNs and Double Deep Q-Networks (DDQNs) for the training of a commercial seven joint redundant manipulator in a real-time trajectory generation task. Experimental results demonstrate that the DDQN approach is more stable then DQN. Moreover, we show that these policies can be directly applied to the official visualizer provided by the robot manufacturer and to the real robot without any further training.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002