paper-with-me

홈 › Papers

Obstruction reasoning for robotic grasping

2025-11-28 · Runyu Jiao, Matteo Bortolon, Francesco Giuliari, Alice Fasoli, Sergio Povoli, Guofeng Mei, Yiming Wang, Fabio Poiesi arxiv

Successful robotic grasping in cluttered environments not only requires a model to visually ground a target object but also to reason about obstructions that must be cleared beforehand. While current vision-language embodied reasoning models show emergent spatial understanding, they remain limited in terms of obstruction reasoning and accessibility planning. To bridge this gap, we present UNOGrasp, a learning-based vision-language model capable of performing visually-grounded obstruction reasoning to infer the sequence of actions needed to unobstruct the path and grasp the target object. We devise a novel multi-step reasoning process based on obstruction paths originated by the target object. We anchor each reasoning step with obstruction-aware visual cues to incentivize reasoning capability. UNOGrasp combines supervised and reinforcement finetuning through verifiable reasoning rewards. Moreover, we construct UNOBench, a large-scale dataset for both training and benchmarking, based on MetaGraspNetV2, with over 100k obstruction paths annotated by humans with obstruction ratios, contact points, and natural-language instructions. Extensive experiments and real-robot evaluations show that UNOGrasp significantly improves obstruction reasoning and grasp success across both synthetic and real-world environments, outperforming generalist and proprietary alternatives. Project website: https://tev-fbk.github.io/UnoGrasp/.

📄 PDF Abstract BibTeX arXiv:2511.23186

Code (0)

등록된 구현이 없습니다.

Tasks

Robotic Grasping

Similar Papers 제목 키워드 기반

AffordanceGrasp-R1:Leveraging Reasoning-Based Affordance Segmentation with Reinforcement Learning for Robotic Grasping

2026-02-03 · Dingyi Zhou, Mu He, Zhuowei Fang, Xiangtong Yao 외 arxiv

We introduce AffordanceGrasp-R1, a reasoning-driven affordance segmentation framework for robotic grasping that combines a chain-of-thought (CoT) cold-start strategy with reinforcement learning to enhance deduction and s…

Reinforcement LearningRobotic Grasping

Beyond Visual Grasping: Benchmarking Complex Grasping from Detection to Execution

2026-07-15 · Hanyi Zhang, Khang Nguyen, Charith Munasinghe, Basu Hela 외 arxiv

Robust robotic grasping remains a fundamental challenge for complex real-world applications. Recent advances in large-scale models demonstrate promising capabilities for reasoning in robotic tasks. However, existing benc…

Robotic Grasping

Free-form language-based robotic reasoning and grasping

2025-03-17 · Runyu Jiao, Alice Fasoli, Francesco Giuliari, Matteo Bortolon 외

Performing robotic grasping from a cluttered bin based on human instructions is a challenging task, as it requires understanding both the nuances of free-form language and the spatial relationships between objects. Visio…

FormRobotic GraspingSpatial ReasoningWorld Knowledge

SECOND-Grasp: Semantic Contact-guided Dexterous Grasping

2026-05-13 · Han Yi Shin, Heeju Ko, Jaewon Mun, Qixing Huang 외 arxiv

Achieving reliable robotic manipulation, such as dexterous grasping, requires a synergy between physically stable interactions and semantic task guidance, yet these objectives are often treated as separate, disjoint goal…

Towards Open-World Grasping with Large Vision-Language Models

2024-06-26 · Georgios Tziafas, Hamidreza Kasaei

The ability to grasp objects in-the-wild from open-ended language instructions constitutes a fundamental challenge in robotics. An open-world grasping system should be able to combine high-level contextual with low-level…

Robotic GraspingVisual GroundingVisual Prompting