paper-with-me

Papers Task Planning

“Task Planning” 태그가 달린 논문 344편 · 필터 해제

Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets

2025-05-21 · Kaiyuan Chen, Shuangyu Xie, Zehan Ma, Pannag R Sanketi 외

Vision-Language Models (VLMs) acquire real-world knowledge and general reasoning ability through Internet-scale image-text corpora. They can augment robotic systems with scene understanding and task planning, and assist …

Dataset GenerationDescriptiveMultiple-choiceQuestion Answering+6

CRAKEN: Cybersecurity LLM Agent with Knowledge-Based Execution

2025-05-21 · Minghao Shao, Haoran Xi, Nanda Rani, Meet Udeshi 외

Large Language Model (LLM) agents can automate cybersecurity tasks and can adapt to the evolving cybersecurity landscape without re-engineering. While LLM agents have demonstrated cybersecurity capabilities on Capture-Th…

Large Language ModelTask PlanningVulnerability Detection

APEX: Empowering LLMs with Physics-Based Task Planning for Real-time Insight

2025-05-20 · Wanjing Huang, Weixiang Yan, Zhen Zhang, Ambuj Singh

Large Language Models (LLMs) demonstrate strong reasoning and task planning capabilities but remain fundamentally limited in physical interaction modeling. Existing approaches integrate perception via Vision-Language Mod…

Causal InferenceDecision Makingmotion predictionReinforcement Learning (RL)+1

Building a Stable Planner: An Extended Finite State Machine Based Planning Module for Mobile GUI Agent

2025-05-20 · Fanglin Mo, Junzhe Chen, Haoxuan Zhu, Xuming Hu

Mobile GUI agents execute user commands by directly interacting with the graphical user interface (GUI) of mobile devices, demonstrating significant potential to enhance user convenience. However, these agents face consi…

Task Planning

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

2025-05-16 · Chenxi Jiang, Chuhao Zhou, Jianfei Yang

Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. Although recent large language model (LLM)-based task planners achieve amazing …

Large Language ModelRobot Task PlanningTask Planning

LODGE: Joint Hierarchical Task Planning and Learning of Domain Models with Grounded Execution

2025-05-15 · Claudius Kienle, Benjamin Alt, Oleg Arenz, Jan Peters

Large Language Models (LLMs) enable planning from natural language instructions using implicit world knowledge, but often produce flawed plans that require refinement. Instead of directly predicting plans, recent methods…

Robot ManipulationTask PlanningWorld Knowledge

Achieving Scalable Robot Autonomy via neurosymbolic planning using lightweight local LLM

2025-05-13 · Nicholas Attolino, Alessio Capitanelli, Fulvio Mastrogiovanni

PDDL-based symbolic task planning remains pivotal for robot autonomy yet struggles with dynamic human-robot collaboration due to scalability, re-planning demands, and delayed plan availability. Although a few neurosymbol…

16k8kTask Planning

PIPA: A Unified Evaluation Protocol for Diagnosing Interactive Planning Agents

2025-05-02 · Takyoung Kim, Janvijay Singh, Shuhaib Mehri, Emre Can Acikgoz 외

The growing capabilities of large language models (LLMs) in instruction-following and context-understanding lead to the era of agents with numerous applications. Among these, task planning agents have become especially p…

Instruction FollowingResponse GenerationTask Planning

CoordField: Coordination Field for Agentic UAV Task Allocation In Low-altitude Urban Scenarios

2025-04-30 · Tengchao Zhang, Yonglin Tian, Fei Lin, Jun Huang 외

With the increasing demand for heterogeneous Unmanned Aerial Vehicle (UAV) swarms to perform complex tasks in urban environments, system design now faces major challenges, including efficient semantic understanding, flex…

Task Planning

LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics

2025-04-30 · Marc Glocker, Peter Hönig, Matthias Hirschmanner, Markus Vincze

We present an embodied robotic system with an LLM-driven agent-orchestration architecture for autonomous household object management. The system integrates memory-augmented task planning, enabling robots to execute high-…

In-Context LearningObjectobject-detectionObject Detection+5

Leveraging Pre-trained Large Language Models with Refined Prompting for Online Task and Motion Planning

2025-04-30 · Huihui Guo, Huilong Pi, Yunchuan Qin, Zhuo Tang 외

With the rapid advancement of artificial intelligence, there is an increasing demand for intelligent robots capable of assisting humans in daily tasks and performing complex operations. Such robots not only require task …

Large Language ModelMotion PlanningTask and Motion PlanningTask Planning

NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks

2025-04-28 · Chia-Yu Hung, Qi Sun, Pengfei Hong, Amir Zadeh 외

Existing Visual-Language-Action (VLA) models have shown promising performance in zero-shot scenarios, demonstrating impressive task execution and reasoning capabilities. However, a significant challenge arises from the l…

Task PlanningVision-Language-ActionVisual Reasoning

Enhancing LLM-Based Agents via Global Planning and Hierarchical Execution

2025-04-23 · Junjie Chen, Haitao Li, Jingli Yang, Yiqun Liu 외

Intelligent agent systems based on Large Language Models (LLMs) have shown great potential in real-world applications. However, existing agent frameworks still face critical limitations in task planning and execution, re…

Task Planning

Robo-Troj: Attacking LLM-based Task Planners

2025-04-23 · Mohaiminul Al Nahian, Zainab Altaweel, David Reitano, Sabbir Ahmed 외

Robots need task planning methods to achieve goals that require more than individual actions. Recently, large language models (LLMs) have demonstrated impressive performance in task planning. LLMs can generate a step-by-…

Backdoor AttackDiversityTask Planning

A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents

2025-04-20 · YuTing Huang, Leilei Ding, Zhipeng Tang, Tianfu Wang 외

Large Language Models (LLMs) exhibit substantial promise in enhancing task-planning capabilities within embodied agents due to their advanced reasoning and comprehension. However, the systemic safety of these agents rema…

BenchmarkingTask Planning

InstructRAG: Leveraging Retrieval-Augmented Generation on Instruction Graphs for LLM-Based Task Planning

2025-04-17 · Zheng Wang, Shu Xian Teo, Jun Jie Chew, Wei Shi

Recent advancements in large language models (LLMs) have enabled their use as agents for planning complex tasks. Existing methods typically rely on a thought-action-observation (TAO) process to enhance LLM performance, b…

Meta-LearningMeta Reinforcement LearningRAGreinforcement-learning+4

FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment

2025-04-11 · Sebastián Barbas Laina, Simon Boche, Sotiris Papatheodorou, Simon Schaefer 외

Geometrically accurate and semantically expressive map representations have proven invaluable to facilitate robust and safe mobile robot navigation and task planning. Nevertheless, real-time, open-vocabulary semantic und…

3D geometryNatural Language QueriesRobot NavigationScene Understanding+1

Personality-Driven Decision-Making in LLM-Based Autonomous Agents

2025-04-01 · Lewis Newsham, Daniel Prince

The embedding of Large Language Models (LLMs) into autonomous agents is a rapidly developing field which enables dynamic, configurable behaviours without the need for extensive domain-specific training. In our previous w…

Decision MakingSchedulingTask Planning

Visual Environment-Interactive Planning for Embodied Complex-Question Answering

2025-04-01 · Ning Lan, Baoshan Ou, Xuemei Xie, Guangming Shi

This study focuses on Embodied Complex-Question Answering task, which means the embodied robot need to understand human questions with intricate structures and abstract semantics. The core of this task lies in making app…

Question AnsweringTask Planning

Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents

2025-04-01 · Saaket Agashe, Kyle Wong, Vincent Tu, Jiachen Yang 외

Computer use agents automate digital tasks by directly interacting with graphical user interfaces (GUIs) on computers and mobile devices, offering significant potential to enhance human productivity by completing an open…

AI AgentTask Planning
← 이전 21–40 / 344 다음 →