paper-with-me

Papers

Diffused Task-Agnostic Milestone Planner

2023-12-06 · NeurIPS 2023 11 · Mineui Hong, Minjae Kang, Songhwai Oh

Addressing decision-making problems using sequence modeling to predict future trajectories shows promising results in recent years. In this paper, we take a step further to leverage the sequence predictive method in wider areas such as long-term planning, vision-based control, and multi-task decision-making. To this end, we propose a method to utilize a diffusion-based generative sequence model to plan a series of milestones in a latent space and to have an agent to follow the milestones to accomplish a given task. The proposed method can learn control-relevant, low-dimensional latent representations of milestones, which makes it possible to efficiently perform long-term planning and vision-based control. Furthermore, our approach exploits generation flexibility of the diffusion model, which makes it possible to plan diverse trajectories for multi-task decision-making. We demonstrate the proposed method across offline reinforcement learning (RL) benchmarks and an visual manipulation environment. The results show that our approach outperforms offline RL methods in solving long-horizon, sparse-reward tasks and multi-task problems, while also achieving the state-of-the-art performance on the most challenging vision-based manipulation benchmark.

📄 PDF Abstract BibTeX arXiv:2312.03395

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingOffline RLReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

One Step at a Time: Long-Horizon Vision-and-Language Navigation with Milestones

2022-02-14 · CVPR 2022 1 · Chan Hee Song, Jihyung Kil, Tai-Yu Pan, Brian M. Sadler 외

We study the problem of developing autonomous agents that can follow human instructions to infer and perform a sequence of actions to complete the underlying task. Significant progress has been made in recent years, espe…

Vision and Language Navigation

Unmasking Interstitial Lung Diseases: Leveraging Masked Autoencoders for Diagnosis

2025-08-06 · Ethan Dack, Lorenzo Brigato, Vasilis Dedousis, Janine Gote-Schniering 외 arxiv

Masked autoencoders (MAEs) have emerged as a powerful approach for pre-training on unlabelled data, capable of learning robust and informative feature representations. This is particularly advantageous in diffused lung d…

MileGPO: Milestone Inference with Local Evidence for Graph-Based Policy Optimization of Long-Horizon LLM Agents

2026-08-20 · Bo Qian, Yuting Wu, Shuang Zeng, Huaiyu Wan 외 arxiv

Credit assignment is challenging in long-horizon agentic reinforcement learning, where supervision often comes only from final rewards. Existing methods refine trajectory-level signals into step-level credits through ste…

Reinforcement Learning

Conversational Disease Diagnosis via External Planner-Controlled Large Language Models

2024-04-04 · Zhoujian Sun, Cheng Luo, Ziyi Liu, Zhengxing Huang

The development of large language models (LLMs) has brought unprecedented possibilities for artificial intelligence (AI) based medical diagnosis. However, the application perspective of LLMs in real diagnostic scenarios …

Active LearningDecision MakingDiagnosticMedical Diagnosis+2

A Novel Narrow Region Detector for Sampling-Based Planners' Efficiency: Match Based Passage Identifier

2025-09-27 · Yafes Enes Şahiner, Esat Yusuf Gündoğdu, Volkan Sezer arxiv

Autonomous technology, which has become widespread today, appears in many different configurations such as mobile robots, manipulators, and drones. One of the most important tasks of these vehicles during autonomous oper…