Real-World Semantic Grasp Detection Based on Attention Mechanism
Recognizing the category of the object and using the features of the object itself to predict grasp configuration is of great significance to improve the accuracy of the grasp detection model and expand its application. Researchers have been trying to combine these capabilities in an end-to-end network to grasping specific objects in a cluttered scene efficiently. In this paper, we propose an end-to-end semantic grasp detection model, which can accomplish both semantic recognition and grasp detection. And we also design a target feature attention mechanism to guide the model focus on the features of target object ontology for grasp prediction according to the semantic information. This method effectively reduces the background features that are weakly correlated to the target object, thus making the features more unique and guaranteeing the accuracy and efficiency of grasp detection. Experimental results show that the proposed method can achieve 98.38% accuracy in Cornell Grasp Dataset. Furthermore, our results on complex multi-object scenarios or more rigorous evaluation metrics show the domain adaptability of our method over the state-of-the-art.
Code (0)
등록된 구현이 없습니다.
Tasks
ObjectSimilar Papers 제목 키워드 기반
Language-driven Grasp Detection with Mask-guided Attention
Grasp detection is an essential task in robotics with various industrial applications. However, traditional methods often struggle with occlusions and do not utilize language for grasping. Incorporating natural language …
Semantic SegmentationBeyond Visual Grasping: Benchmarking Complex Grasping from Detection to Execution
Robust robotic grasping remains a fundamental challenge for complex real-world applications. Recent advances in large-scale models demonstrate promising capabilities for reasoning in robotic tasks. However, existing benc…
Robotic GraspingWhen Transformer Meets Robotic Grasping: Exploits Context for Efficient Grasp Detection
In this paper, we present a transformer-based architecture, namely TF-Grasp, for robotic grasp detection. The developed TF-Grasp framework has two elaborate designs making it well suitable for visual grasping tasks. The …
DecoderRobotic GraspingMulti-fingered Robotic Hand Grasping in Cluttered Environments through Hand-object Contact Semantic Mapping
The deep learning models has significantly advanced dexterous manipulation techniques for multi-fingered hand grasping. However, the contact information-guided grasping in cluttered environments remains largely underexpl…
Dataset GenerationDiversityGrasp GenerationNBMOD: Find It and Grasp It in Noisy Background
Grasping objects is a fundamental yet important capability of robots, and many tasks such as sorting and picking rely on this skill. The prerequisite for stable grasping is the ability to correctly identify suitable gras…
Robotic Grasping