paper-with-me

홈 › Papers

End-to-End Learning of Semantic Grasping

2017-07-06 · Eric Jang, Sudheendra Vijayanarasimhan, Peter Pastor, Julian Ibarz, Sergey Levine

We consider the task of semantic robotic grasping, in which a robot picks up an object of a user-specified class using only monocular images. Inspired by the two-stream hypothesis of visual reasoning, we present a semantic grasping framework that learns object detection, classification, and grasp planning in an end-to-end fashion. A "ventral stream" recognizes object class while a "dorsal stream" simultaneously interprets the geometric relationships necessary to execute successful grasps. We leverage the autonomous data collection capabilities of robots to obtain a large self-supervised dataset for training the dorsal stream, and use semi-supervised label propagation to train the ventral stream with only a modest amount of human supervision. We experimentally show that our approach improves upon grasping systems whose components are not learned end-to-end, including a baseline method that uses bounding box detection. Furthermore, we show that jointly training our model with auxiliary data consisting of non-semantic grasping data, as well as semantically labeled images without grasp actions, has the potential to substantially improve semantic grasping performance.

📄 PDF Abstract BibTeX arXiv:1707.01932

Code (0)

등록된 구현이 없습니다.

Tasks

Objectobject-detectionObject DetectionRobotic GraspingVisual Reasoning

Similar Papers 제목 키워드 기반

Multi-fingered Robotic Hand Grasping in Cluttered Environments through Hand-object Contact Semantic Mapping

2024-04-12 · Lei Zhang, Kaixin Bai, Guowen Huang, Zhenshan Bing 외

The deep learning models has significantly advanced dexterous manipulation techniques for multi-fingered hand grasping. However, the contact information-guided grasping in cluttered environments remains largely underexpl…

Dataset GenerationDiversityGrasp Generation

SECOND-Grasp: Semantic Contact-guided Dexterous Grasping

2026-05-13 · Han Yi Shin, Heeju Ko, Jaewon Mun, Qixing Huang 외 arxiv

Achieving reliable robotic manipulation, such as dexterous grasping, requires a synergy between physically stable interactions and semantic task guidance, yet these objectives are often treated as separate, disjoint goal…

Beyond Visual Grasping: Benchmarking Complex Grasping from Detection to Execution

2026-07-15 · Hanyi Zhang, Khang Nguyen, Charith Munasinghe, Basu Hela 외 arxiv

Robust robotic grasping remains a fundamental challenge for complex real-world applications. Recent advances in large-scale models demonstrate promising capabilities for reasoning in robotic tasks. However, existing benc…

Robotic Grasping

FineGrasp: Towards Robust Grasping for Delicate Objects

2025-07-08 · Yun Du, Mengao Zhao, Tianwei Lin, Yiwei Jin 외 arxiv

Recent advancements in robotic grasping have led to its integration as a core module in many manipulation systems. For instance, language-driven semantic segmentation enables the grasping of any designated object or obje…

Semantic SegmentationRobotic Grasping

IFG: Internet-Scale Guidance for Functional Grasping Generation

2025-11-12 · Ray Muxin Liu, Mingxuan Li, Kenneth Shaw, Deepak Pathak arxiv

Large Vision Models trained on internet-scale data have demonstrated strong capabilities in segmenting and semantically understanding object parts, even in cluttered, crowded scenes. However, while these models can direc…

Point Clouds