paper-with-me

홈 › Papers

Language-Enhanced Mobile Manipulation for Efficient Object Search in Indoor Environments

2025-08-28 · Liding Zhang, Zeqi Li, Kuanqi Cai, Qian Huang, Zhenshan Bing, Alois Knoll arxiv

Enabling robots to efficiently search for and identify objects in complex, unstructured environments is critical for diverse applications ranging from household assistance to industrial automation. However, traditional scene representations typically capture only static semantics and lack interpretable contextual reasoning, limiting their ability to guide object search in completely unfamiliar settings. To address this challenge, we propose a language-enhanced hierarchical navigation framework that tightly integrates semantic perception and spatial reasoning. Our method, Goal-Oriented Dynamically Heuristic-Guided Hierarchical Search (GODHS), leverages large language models (LLMs) to infer scene semantics and guide the search process through a multi-level decision hierarchy. Reliability in reasoning is achieved through the use of structured prompts and logical constraints applied at each stage of the hierarchy. For the specific challenges of mobile manipulation, we introduce a heuristic-based motion planner that combines polar angle sorting with distance prioritization to efficiently generate exploration paths. Comprehensive evaluations in Isaac Sim demonstrate the feasibility of our framework, showing that GODHS can locate target objects with higher search efficiency compared to conventional, non-semantic search strategies. Website and Video are available at: https://drapandiger.github.io/GODHS

📄 PDF Abstract BibTeX arXiv:2508.20899

Code (0)

등록된 구현이 없습니다.

Tasks

Spatial Reasoning

Similar Papers 제목 키워드 기반

MobileManiBench: Simplifying Model Verification for Mobile Manipulation

2026-02-05 · Wenbo Wang, Fangyun Wei, QiXiu Li, Xi Chen 외 arxiv

Vision-language-action models have advanced robotic manipulation but remain constrained by reliance on the large, teleoperation-collected datasets dominated by the static, tabletop scenes. We propose a simulation-first f…

Reinforcement Learning

MobileVLA-R1 2.0: RL-Enhanced Reasoning for Mobile Robot Control

2026-09-05 · Ting Huang, Yue Huang, Zeyu Zhang, Shuicheng Yan 외 hf

Grounding natural-language instructions into reliable and executable actions remains a fundamental challenge for vision-language-action (VLA) systems on mobile robots, due to the persistent gap between high-level semanti…

Reinforcement LearningInstruction FollowingMultimodal ReasoningDecision Making

RoboMIND 2.0: A Multimodal, Bimanual Mobile Manipulation Dataset for Generalizable Embodied Intelligence

2025-12-31 · Chengkai Hou, Kun Wu, Jiaming Liu, Zhengping Che 외 arxiv

While data-driven imitation learning has revolutionized robotic manipulation, current approaches remain constrained by the scarcity of large-scale, diverse real-world demonstrations. Consequently, the ability of existing…

Reinforcement Learning

UniTeam: Open Vocabulary Mobile Manipulation Challenge

2023-12-14 · Andrew Melnik, Michael Büttner, Leon Harz, Lyon Brown 외

This report introduces our UniTeam agent - an improved baseline for the "HomeRobot: Open Vocabulary Mobile Manipulation" challenge. The challenge poses problems of navigation in unfamiliar environments, manipulation of n…

Object

MoDeSuite: Robot Learning Task Suite for Benchmarking Mobile Manipulation with Deformable Objects

2025-07-29 · Yuying Zhang, Kevin Sebastian Luck, Francesco Verdoja, Ville Kyrki 외 arxiv

Mobile manipulation is a critical capability for robots operating in diverse, real-world environments. However, manipulating deformable objects and materials remains a major challenge for existing robot learning algorith…

Reinforcement Learning