Bimodal SegNet: Instance Segmentation Fusing Events and RGB Frames for Robotic Grasping
Object segmentation for robotic grasping under dynamic conditions often faces challenges such as occlusion, low light conditions, motion blur and object size variance. To address these challenges, we propose a Deep Learning network that fuses two types of visual signals, event-based data and RGB frame data. The proposed Bimodal SegNet network has two distinct encoders, one for each signal input and a spatial pyramidal pooling with atrous convolutions. Encoders capture rich contextual information by pooling the concatenated features at different resolutions while the decoder obtains sharp object boundaries. The evaluation of the proposed method undertakes five unique image degradation challenges including occlusion, blur, brightness, trajectory and scale variance on the Event-based Segmentation (ESD) Dataset. The evaluation results show a 6-10\% segmentation accuracy improvement over state-of-the-art methods in terms of mean intersection over the union and pixel accuracy. The model code is available at https://github.com/sanket0707/Bimodal-SegNet.git
Code (1)
Tasks
DecoderInstance SegmentationObjectRobotic GraspingSegmentationSemantic SegmentationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
U-NetMN and SegNetMN: Modified U-Net and SegNet models for bimodal SAR image segmentation
Segmenting Synthetic Aperture Radar (SAR) images is crucial for many remote sensing applications, particularly water body detection. However, deep learning-based segmentation models often face challenges related to conve…
Body DetectionComputational EfficiencyImage SegmentationSegmentation+1SegNet4D: Efficient Instance-Aware 4D Semantic Segmentation for LiDAR Point Cloud
4D LiDAR semantic segmentation, also referred to as multi-scan semantic segmentation, plays a crucial role in enhancing the environmental understanding capabilities of autonomous vehicles or robots. It classifies the sem…
Autonomous DrivingAutonomous NavigationAutonomous VehiclesLIDAR Semantic Segmentation+2FUSegNet: A Deep Convolutional Neural Network for Foot Ulcer Segmentation
This paper presents FUSegNet, a new model for foot ulcer segmentation in diabetes patients, which uses the pre-trained EfficientNet-b7 as a backbone to address the issue of limited training samples. A modified spatial an…
DecoderMultiple Instance Segmentation in Brachial Plexus Ultrasound Image Using BPMSegNet
The identification of nerve is difficult as structures of nerves are challenging to image and to detect in ultrasound images. Nevertheless, the nerve identification in ultrasound images is a crucial step to improve perfo…
Instance SegmentationSemantic SegmentationEV-LayerSegNet: Self-supervised Motion Segmentation using Event Cameras
Event cameras are novel bio-inspired sensors that capture motion dynamics with much higher temporal resolution than traditional cameras, since pixels react asynchronously to brightness changes. They are therefore better …
DeblurringMotion SegmentationOptical Flow EstimationSegmentation+1