首页 /研究 /Bimodal SegNet: Fused instance segmentation using events and RGB frames
MANIPULATION

Bimodal SegNet: Fused instance segmentation using events and RGB frames

Sanket Kachole, Xiaoqian Huang, Fariborz Baghaei Naeini, Rajkumar Muthusamy, Dimitrios Makris, Yahya Zweiri

发表年份
2023
引用次数
12

摘要

Object segmentation enhances robotic grasping by aiding object identification. Complex environments and dynamic conditions pose challenges such as occlusion, low light conditions, motion blur and object size variance. To address these challenges, we propose a Bimodal SegNet that fuses two types of visual signals, event-based data and RGB frame data. The proposed Bimodal SegNet network has two distinct encoders — one for RGB signal input and another for Event signal input, in addition to an Atrous Pyramidal Feature Amplification module. Encoders capture and fuse the rich contextual information from different resolutions via a Cross-Domain Contextual Attention layer while the decoder obtains sharp object boundaries. The evaluation of the proposed method undertakes five unique image degradation challenges including occlusion, blur, brightness, trajectory and scale variance on the Event-based Segmentation (ESD) Dataset. The results show a 4%–6% MIOU score improvement over state-of-the-art methods in terms of mean intersection over the union and pixel accuracy. The source code, dataset and model are publicly available at: https://github.com/sanket0707/Bimodal-SegNet . • Bimodal SegNet: A novel encoder-decoder with transformer cross-attention. • Dual-encoder for bimodal input, merging features with weighted cross-attention. • Tested on ESD dataset, Bimodal SegNet excels in mIoU and Pixel accuracy.

关键词

Artificial intelligenceComputer scienceComputer visionRGB color modelSegmentationEncoderPattern recognition (psychology)Feature (linguistics)Motion blurEvent (particle physics)

相关论文

查看 MANIPULATION 分类全部论文