首页 /研究 /Policy Fusion Transfer: The Knowledge Transfer for Different Robot Peg-in-Hole Insertion Assemblies
LEARNING

Policy Fusion Transfer: The Knowledge Transfer for Different Robot Peg-in-Hole Insertion Assemblies

Yu Men, Ligang Jin, Tao Cui, Yunfeng Bai, Fengming Li, Rui Song

发表年份
2023
引用次数
16

摘要

Given the problem of exploration and utilization, the high interaction cost of deep reinforcement learning is sometimes unacceptable. Moreover, the model trained by deep reinforcement learning cannot adapt to the complex and changeable assembly environment quickly owing to the weak generalization ability. Thus, the knowledge transfer framework is established for different robot peg-in-hole assemblies in this article to reinforce the generalization ability of the assembly model and the utilization efficiency of data. In the framework, the source domain model is trained through the Proximal Policy Optimization algorithm. The model accurately predicts environmental information and dynamically adjusts the robot’s movements based on the prediction. Combined with the PI controller, the flexible peg-in-hole assembly can be quickly achieved by the robot. The policy fusion between the source and target domain MDP is realized by the Policy Fusion Transfer algorithm designed in this article. The effectiveness of the algorithm is verified and discussed based on the transfer experiments, involving different assembly objects and different assembly environments. The training results demonstrate that the environment is explored by the Policy Fusion Transfer algorithm in a safe range at the initial stage of training by reusing the assembly policy of the source domain and that the faster model convergence speed is achieved. Furthermore, the testing results suggest that the assembly success rate is improved by 26% and the assembly force is reduced by 30N. Moreover, the knowledge transfer between different peg-in-hole assemblies is realized successfully.

关键词

GeneralizationComputer scienceReinforcement learningRobotDomain (mathematical analysis)Transfer of learningConvergence (economics)Controller (irrigation)FusionTransfer (computing)

相关论文

查看 LEARNING 分类全部论文