paper-with-me

홈 › Papers

Hindsight Planner: A Closed-Loop Few-Shot Planner for Embodied Instruction Following

2024-12-27 · Yuxiao Yang, Shenao Zhang, Zhihan Liu, Huaxiu Yao, Zhaoran Wang

This work focuses on building a task planner for Embodied Instruction Following (EIF) using Large Language Models (LLMs). Previous works typically train a planner to imitate expert trajectories, treating this as a supervised task. While these methods achieve competitive performance, they often lack sufficient robustness. When a suboptimal action is taken, the planner may encounter an out-of-distribution state, which can lead to task failure. In contrast, we frame the task as a Partially Observable Markov Decision Process (POMDP) and aim to develop a robust planner under a few-shot assumption. Thus, we propose a closed-loop planner with an adaptation module and a novel hindsight method, aiming to use as much information as possible to assist the planner. Our experiments on the ALFRED dataset indicate that our planner achieves competitive performance under a few-shot assumption. For the first time, our few-shot agent's performance approaches and even surpasses that of the full-shot supervised agent.

📄 PDF Abstract BibTeX arXiv:2412.19562

Code (0)

등록된 구현이 없습니다.

Tasks

Instruction Following

Similar Papers 제목 키워드 기반

Can Vehicle Motion Planning Generalize to Realistic Long-tail Scenarios?

2024-04-11 · Marcel Hallgarten, Julian Zapata, Martin Stoll, Katrin Renz 외

Real-world autonomous driving systems must make safe decisions in the face of rare and diverse traffic scenarios. Current state-of-the-art planners are mostly evaluated on real-world datasets like nuScenes (open-loop) or…

Autonomous DrivingMotion PlanningNavigate

When Planners Meet Reality: How Learned, Reactive Traffic Agents Shift nuPlan Benchmarks

2025-10-16 · Steffen Hagedorn, Luka Donkov, Aron Distelzweig, Alexandru P. Condurache arxiv

Planner evaluation in closed-loop simulation often uses rule-based traffic agents, whose simplistic and passive behavior can hide planner deficiencies and bias rankings. Widely used IDM agents simply follow a lead vehicl…

Recursively Feasible Chance-constrained Model Predictive Control under Gaussian Mixture Model Uncertainty

2024-01-08 · Kai Ren, Colin Chen, Hyeontae Sung, Heejin Ahn 외

We present a chance-constrained model predictive control (MPC) framework under Gaussian mixture model (GMM) uncertainty. Specifically, we consider the uncertainty that arises from predicting future behaviors of moving ob…

Autonomous DrivingmodelModel Predictive ControlTrajectory Prediction

PlannerRFT: Reinforcing Diffusion Planners through Closed-Loop and Sample-Efficient Fine-Tuning

2026-01-19 · Hongchen Li, Tianyu Li, Jiazhi Yang, Haochen Tian 외 arxiv

Diffusion-based planners have emerged as a promising approach for human-like trajectory generation in autonomous driving. Recent works incorporate reinforcement fine-tuning to enhance the robustness of diffusion planners…

Autonomous Driving

Shift & Drift: A Zero-Shot Benchmark for Generalizable and Robust Autonomous Driving Motion Planning

2026-07-08 · Alessandro Canevaro, Hang Yu, Julian Schmidt, Peizheng Li 외 arxiv

While closed-loop motion planners trained on large-scale, object-level datasets, e.g., nuPlan, demonstrate strong in-distribution (ID) performance, their generalization to novel urban topologies and recovery mechanisms f…

Autonomous DrivingMotion Planning