首页 /研究 /Reinforcement Learning based Path Planning of the Mobile Agents with Constrained Locomotion for the Material Handling Applications
SWARM

Reinforcement Learning based Path Planning of the Mobile Agents with Constrained Locomotion for the Material Handling Applications

Satheeshkumar Veeramani, Sreekumar Muthuswamy

发表年份
2020
引用次数
3

摘要

This paper presents the intelligent path planning model of the mobile base agent of SwarmItFIX robot with novel Swing and Dock (SaD) locomotion for material handling/transfer applications. In this work, the Markov Decision Process (MDP) path planning problem of SaD agent is solved using two Reinforcement Learning (RL) based dynamic programming methods viz Policy Iteration (PI), and Value Iteration (VI). Being tested with 16 different test cases, both the algorithms return the optimal sequence of steps with reduced makespan for the mobile agents to reach the goal positions positively. The results of both the methods were compared with each other in the section V, and found to be convincing. Hence the proposed control scheme is being implemented in the SwarmItFIX setup available at the University of Genova, Italy.

关键词

Reinforcement learningMarkov decision processMotion planningComputer scienceMobile robotQ-learningMathematical optimizationPath (computing)Dynamic programmingPartially observable Markov decision process

相关论文

查看 SWARM 分类全部论文