首页 /研究 /MT-DSSD: multi-task deconvolutional single shot detector for object detection, segmentation, and grasping detection
MANIPULATION

MT-DSSD: multi-task deconvolutional single shot detector for object detection, segmentation, and grasping detection

Ryosuke Araki, Tsubasa Hirakawa, Takayoshi Yamashita, Hironobu Fujiyoshi

发表年份
2022
引用次数
16

摘要

A robot that picks and places the wide variety of items in a logistics warehouse must detect and recognize items from images and then decide which points to grasp. Our Multi-task Deconvolutional Single Shot Detector (MT-DSSD) simultaneously performs the three tasks necessary for this manipulation: object detection, semantic segmentation, and grasping detection. MT-DSSD is a multi-task learning (MTL) method based on DSSD that reduces the amount of computation and achieves high speed compared to when separate models perform each task. Evaluations using the Amazon Robotics Challenge dataset showed that our model has a better object detection and segmentation performance than comparable methods, and an ablation study showed that MTL could improve the accuracy of each task. Further, robotic experiments for grasping demonstrated that our model could detect the appropriate grasping point.

关键词

Artificial intelligenceComputer scienceComputer visionSegmentationObject detectionTask (project management)DetectorRoboticsGRASPRobot

相关论文

查看 MANIPULATION 分类全部论文