paper-with-me

홈 › Papers

Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control

2026-07-12 · Jun Chen, Erdent Bao, Wenlong Dong, Jierui Liu, Qi Cai, Hao Wan, Shaopeng Li, Weijun Qin, Jing Liang, Huiping Zhuang arxiv

Language-conditioned Imitation Learning (IL) is essential for enabling robots to perform complex tasks following natural language instructions. However, generalizing to multi-step compositional tasks remains a significant challenge. While hierarchical approaches attempt to address this by decomposing tasks into atomic skills, existing methods often suffer from training instability and codebook collapse due to the tight coupling between high-level skill reasoning and low-level action generation in joint training paradigms. Inspired by the Dual-Process Theory of cognition, we propose Dual-Process Atomic Skill Learning (DASL), a novel asynchronous hierarchical imitation learning framework that decouples slow semantic reasoning from fast, real-time motion control. DASL comprises a Slow-Frequency Policy that predicts interpretable, discrete skills via Vector Quantization, and a High-Frequency Policy that leverages a latent diffusion model and a Decision Transformer to generate precise actions conditioned on these latent skills. By asynchronously coordinating these modules and utilizing diffusion to structure the latent space, our framework mitigates the skill codebook interference problem common in joint training paradigms. Evaluations across simulation benchmarks and experiment demonstrate that DASL significantly outperforms state-of-the-art baselines, excelling in skill acquisition and compositional generalization to unseen instructions. GitHub page: https://github.com/Hatakekaka/DASL

📄 PDF Abstract BibTeX arXiv:2607.10625

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation

2025-12-20 · Yihang Zhu, Weiqing Wang, Shijie Wu, Ye Shi 외 arxiv

Scaling imitation learning to diverse multi-task robot manipulation remains challenging due to suboptimal demonstrations, behavioral multi-modality, and destructive interference across tasks. While skill-based methods of…

Robot Manipulation

RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics

2025-04-02 · Zhiyuan Zhang, Yuxin He, Yong Sun, Junyu Shi 외

Visual Language Models (VLMs) have emerged as pivotal tools for robotic systems, enabling cross-task generalization, dynamic environmental interaction, and long-horizon planning through multimodal perception and semantic…

Action UnderstandingRepresentation Learning

Atomic Skills are the Prerequisite: When Reinforcement Learning Synthesizes Compositional Reasoning, and When It Only Amplifies

2025-12-01 · Sitao Cheng, Xunjian Yin, Ruiwen Zhou, Yuxuan Li 외 arxiv

Does Reinforcement Learning (RL) merely amplify existing skills, or synthesize novel skills? We investigate this question through the lens of Complementary Reasoning: the critical practical capability of integrating inte…

Reinforcement LearningContinual Learning

Task-Differentiated Atomic Skill Expansion and Routing for Continual Learning Across Highly Heterogeneous Tasks

2026-06-19 · Jiacheng Wang, Xinjia He, Qi Ding, Yutao Yang 외 arxiv

Continual learning (CL) is commonly studied under the assumption that sequential tasks are semantically related or structurally similar. However, in highly heterogeneous settings, where tasks differ substantially in reas…

Incremental LearningContinual Learning

ProSAC-CT: Progressive Spectral-Anatomical Co-Guided Multi-Stage Diffusion Model for Low-Dose CT Denoising

2026-07-02 · Xuepeng Liu, Zetong Liu, Renyiming Li, Yan Li 외 arxiv

Low-dose computed tomography (LDCT) reduces radiation exposure but introduces stronger quantum noise, streak artifacts, and local texture degradation, which can obscure anatomical boundaries and weaken low-contrast struc…