paper-with-me

홈 › Papers

Uncertainty Comes for Free: Human-in-the-Loop Policies with Diffusion Models

2025-02-26 · Zhanpeng He, Yifeng Cao, Matei Ciocarlie

Human-in-the-loop (HitL) robot deployment has gained significant attention in both academia and industry as a semi-autonomous paradigm that enables human operators to intervene and adjust robot behaviors at deployment time, improving success rates. However, continuous human monitoring and intervention can be highly labor-intensive and impractical when deploying a large number of robots. To address this limitation, we propose a method that allows diffusion policies to actively seek human assistance only when necessary, reducing reliance on constant human oversight. To achieve this, we leverage the generative process of diffusion policies to compute an uncertainty-based metric based on which the autonomous agent can decide to request operator assistance at deployment time, without requiring any operator interaction during training. Additionally, we show that the same method can be used for efficient data collection for fine-tuning diffusion policies in order to improve their autonomous performance. Experimental results from simulated and real-world environments demonstrate that our approach enhances policy performance during deployment for a variety of scenarios.

📄 PDF Abstract BibTeX arXiv:2503.01876

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Distribution-Free Uncertainty Quantification in Mechanical Ventilation Treatment: A Conformal Deep Q-Learning Framework

2024-12-17 · Niloufar Eghbali, Tuka Alhanai, Mohammad M. Ghassemi

Mechanical Ventilation (MV) is a critical life-support intervention in intensive care units (ICUs). However, optimal ventilator settings are challenging to determine because of the complexity of balancing patient-specifi…

Conformal PredictionDeep Reinforcement LearningQ-LearningUncertainty Quantification

Learning Visuotactile Estimation and Control for Non-prehensile Manipulation under Occlusions

2024-12-17 · Juan Del Aguila Ferrandis, João Moura, Sethu Vijayakumar

Manipulation without grasping, known as non-prehensile manipulation, is essential for dexterous robots in contact-rich environments, but presents many challenges relating with underactuation, hybrid-dynamics, and frictio…

Reinforcement Learning (RL)

Hand-in-the-Loop: Improving VLA Policies for Dexterous Manipulation via Seamless Hand-Arm Intervention

2026-05-14 · Zhuohang Li, Liqun Huang, Wei Xu, Zhengming Zhu 외 arxiv

Vision-Language-Action (VLA) models are prone to compounding errors in dexterous manipulation, where high-dimensional action spaces and contact-rich dynamics amplify small policy deviations over long horizons. While Inte…

A Human-in-the-Loop Confidence-Aware Failure Recovery Framework for Modular Robot Policies

2026-02-10 · Rohan Banerjee, Krishna Palempalli, Bohan Yang, Jiaying Fang 외 arxiv

Robots operating in unstructured human environments inevitably encounter failures, especially in robot caregiving scenarios. While humans can often help robots recover, excessive or poorly targeted queries impose unneces…

LLM-Augmented Digital Twin for Policy Evaluation in Short-Video Platforms

2026-03-11 · Haoting Zhang, Yunduan Lin, Jinghai He, Denglin Jiang 외 arxiv

Short-video platforms are closed-loop, human-in-the-loop ecosystems where platform policy, creator incentives, and user behavior co-evolve. This feedback structure makes counterfactual policy evaluation difficult in prod…