paper-with-me

Papers

Towards Embodied Agentic AI: Review and Classification of LLM- and VLM-Driven Robot Autonomy and Interaction

2025-08-07 · Sahar Salimpour, Lei Fu, Kajetan Rachwał, Pascal Bertrand, Kevin O'Sullivan, Robert Jakob, Farhad Keramat, Leonardo Militano, Giovanni Toffetti, Harry Edelman, Jorge Peña Queralta arxiv

Foundation models, including large language models (LLMs) and vision-language models (VLMs), have recently enabled novel approaches to robot autonomy and human-robot interfaces. In parallel, vision-language-action models (VLAs) or large behavior models (LBMs) are increasing the dexterity and capabilities of robotic systems. This survey paper reviews works that advance agentic applications and architectures, including initial efforts with GPT-style interfaces and more complex systems where AI agents function as coordinators, planners, perception actors, or generalist interfaces. Such agentic architectures allow robots to reason over natural language instructions, invoke APIs, plan task sequences, or assist in operations and diagnostics. In addition to peer-reviewed research, due to the fast-evolving nature of the field, we highlight and include community-driven projects, ROS packages, and industrial frameworks that show emerging trends. We propose a taxonomy for classifying model integration approaches and present a comparative analysis of the role that agents play in different solutions in today's literature.

📄 PDF Abstract BibTeX arXiv:2508.05294

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When Multi-Robot Systems Meet Agentic AI:Towards Embodied Collective Intelligence

2026-06-26 · Yuxuan Yan, Yuanyuan Jia, Qianqian Yang arxiv

Embodied AI is increasingly becoming agentic, shifting robots from perception--control pipelines towards closed-loop systems that can retrieve context, deliberate during execution, monitor feedback, and refine future beh…

Vision-Language-Action Models: Concepts, Progress, Applications and Challenges

2025-05-07 · Ranjan Sapkota, Yang Cao, Konstantinos I. Roumeliotis, Manoj Karkee

Vision-Language-Action (VLA) models mark a transformative advancement in artificial intelligence, aiming to unify perception, natural language understanding, and embodied action within a single computational framework. T…

Autonomous VehiclesNatural Language UnderstandingVision-Language-Action

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses

2026-03-28 · Xiao Li, Xiang Zheng, Yifeng Gao, Xinyu Xia 외 arxiv

Embodied Artificial Intelligence (Embodied AI) integrates perception, cognition, planning, and interaction into agents that operate in open-world, safety-critical environments. As these systems gain autonomy and enter do…

3D Generation for Embodied AI and Robotic Simulation: A Survey

2026-04-29 · Tianwei Ye, Yifan Mao, Minwen Liao, Jian Liu 외 arxiv

Embodied AI and robotic systems increasingly depend on scalable, diverse, and physically grounded 3D content for simulation-based training and real-world deployment. While 3D generative modeling has advanced rapidly, emb…

Data AugmentationScene Generation3D Generation

EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents

2026-04-20 · Paolo Riva, Leonardo Gargani, Matteo Frosi, Matteo Matteucci arxiv

As the world of agentic artificial intelligence applied to robotics evolves, the need for agents capable of building and retrieving memories and observations efficiently is increasing. Robots operating in complex environ…