paper-with-me

홈 › Papers

Grounding Dynamic Spatial Relations for Embodied (Robot) Interaction

2016-07-26 · Michael Spranger, Jakob Suchan, Mehul Bhatt, Manfred Eppe

This paper presents a computational model of the processing of dynamic spatial relations occurring in an embodied robotic interaction setup. A complete system is introduced that allows autonomous robots to produce and interpret dynamic spatial phrases (in English) given an environment of moving objects. The model unites two separate research strands: computational cognitive semantics and on commonsense spatial representation and reasoning. The model for the first time demonstrates an integration of these different strands.

📄 PDF Abstract BibTeX arXiv:1607.07565

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EmbodimentSemantic: A Spatial Scene-Graph Dataset and Benchmark for Vision-Language Models on Embodied Manipulation Trajectories

2026-06-06 · Hassan Jaber, Refinath S N, Luca Cagliero, Christopher E. Mower 외 arxiv

Spatial grounding remains a key limitation of vision-language-action (VLA) systems for robotic manipulation. While current models can recognize objects and follow language instructions, they often lack an explicit repres…

Searching in Space and Time: Unified Memory-Action Loops for Open-World Object Retrieval

2025-11-18 · Taijing Chen, Sateesh Kumar, Junhong Xu, Georgios Pavlakos 외 arxiv

Service robots must retrieve objects in dynamic, open-world settings where requests may reference attributes ("the red mug"), spatial context ("the mug on the table"), or past states ("the mug that was here yesterday"). …

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI

2026-07-05 · ACE-Brain Team, :, Ziyang Gong, Haoming Gu 외 arxiv

Embodied AI is moving from isolated perception or action modules toward physical agents that understand, plan under goals, act through robot bodies, monitor progress, and improve from experience. Existing systems address…

Spatial ReasoningDecision Making

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

2026-06-22 · Xiaolin Zhou, Liu Liu, Tingyang Xiao, Wei Feng 외 arxiv

LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise actions. Extending this loop to physical robots is difficult because ph…

Pointing-VLA: Typed Spatial Grounding Interfaces for Vision-Language-Action Manipulation

2026-08-24 · Xiwen Chen, Zelin Li, Zhiruo Zhou, Huiming Chen 외 arxiv

Vision-language-action (VLA) models often expose spatial grounding through autoregressive text coordinates or opaque action tokens, creating brittle interfaces between multimodal reasoning and robot execution. We present…

Multimodal Reasoning