paper-with-me

Papers

Flexible Multitask Learning with Factorized Diffusion Policy

2025-12-26 · Chaoqi Liu, Haonan Chen, Sigmund H. Høeg, Shaoxiong Yao, Yunzhu Li, Kris Hauser, Yilun Du arxiv

Multitask learning poses significant challenges due to the highly multimodal and diverse nature of robot action distributions. However, effectively fitting policies to these complex task distributions is often difficult, and existing monolithic models often underfit the action distribution and lack the flexibility required for efficient adaptation. We introduce a novel modular diffusion policy framework that factorizes complex action distributions into a composition of specialized diffusion models, each capturing a distinct sub-mode of the behavior space for a more effective overall policy. In addition, this modular structure enables flexible policy adaptation to new tasks by adding or fine-tuning components, which inherently mitigates catastrophic forgetting. Empirically, across both simulation and real-world robotic manipulation settings, we illustrate how our method consistently outperforms strong modular and monolithic baselines.

📄 PDF Abstract BibTeX arXiv:2512.21898

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning

2024-07-01 · Yixiao Wang, Yifei Zhang, Mingxiao Huo, Ran Tian 외

The increasing complexity of tasks in robotics demands efficient strategies for multitask and continual learning. Traditional models typically rely on a universal policy for all tasks, facing challenges such as high comp…

Continual LearningMixture-of-Experts

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

2025-11-17 · Jonas Bode, Raphael Memmesheimer, Sven Behnke arxiv

Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and its application to physical tasks. This paper seeks to harness the capabilitie…

Image Generation

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

2026-05-24 · Lei Zheng, Peiqi Yu, Zengqi Peng, Changliu Liu 외 arxiv

Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency that is incompatible with high-frequency closed-loop control. To add…

ECHO: Efficient Chest X-ray Report Generation with One-step Block Diffusion

2026-04-10 · Lifeng Chen, Tianqi You, Hao Liu, Zhimin Bao 외 arxiv

Chest X-ray report generation (CXR-RG) has the potential to substantially alleviate radiologists' workload. However, conventional autoregressive vision--language models (VLMs) suffer from high inference latency due to se…

Factorizing Diffusion Policies for Observation Modality Prioritization

2025-09-20 · Omkar Patil, Prabin Rath, Kartikay Pangaonkar, Eric Rosen 외 arxiv

Diffusion models have been extensively leveraged for learning robot skills from demonstrations. These policies are conditioned on several observational modalities such as proprioception, vision and tactile. However, obse…