paper-with-me

홈 › Papers

Lifelong Language-Conditioned Robotic Manipulation Learning

2026-03-05 · Xudong Wang, Zebin Han, Zhiyu Liu, Gan Li, Jiahua Dong, Baichen Liu, Lianqing Liu, Zhi Han arxiv

Traditional language-conditioned manipulation agent sequential adaptation to new manipulation skills leads to catastrophic forgetting of old skills, limiting dynamic scene practical deployment. In this paper, we propose SkillsCrafter, a novel robotic manipulation framework designed to continually learn multiple skills while reducing catastrophic forgetting of old skills. Specifically, we propose a Manipulation Skills Adaptation to retain the old skills knowledge while inheriting the shared knowledge between new and old skills to facilitate learning of new skills. Meanwhile, we perform the singular value decomposition on the diverse skill instructions to obtain common skill semantic subspace projection matrices, thereby recording the essential semantic space of skills. To achieve forget-less and generalization manipulation, we propose a Skills Specialization Aggregation to compute inter-skills similarity in skill semantic subspaces, achieving aggregation of the previously learned skill knowledge for any new or unknown skill. Extensive experiments demonstrate the effectiveness and superiority of our proposed SkillsCrafter.

📄 PDF Abstract BibTeX arXiv:2603.05160

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Human-like Physical Intelligence: LifelongVision-Language-Action Learning for Robotic Manipulation

2026-07-16 · Yao He, Gan Sun, Wenqi Liang, Fazeng Li 외 arxiv

Similar to the natural capabilities of humans to sequentially learn new tasks, robots with Vision-Language-Action (VLA) models should possess lifelong learning ability to learn a new task when deployed in open-world envi…

CRL-VLA: Continual Vision-Language-Action Learning

2026-02-03 · Qixin Zeng, Shuo Zhang, Hongyin Zhang, Renjie Wang 외 arxiv

Lifelong learning is critical for embodied agents in open-world environments, where reinforcement learning fine-tuning has emerged as an important paradigm to enable Vision-Language-Action (VLA) models to master dexterou…

Reinforcement Learning

GVCCI: Lifelong Learning of Visual Grounding for Language-Guided Robotic Manipulation

2023-07-12 · Junghyun Kim, Gi-Cheon Kang, Jaein Kim, Suyeon Shin 외

Language-Guided Robotic Manipulation (LGRM) is a challenging task as it requires a robot to understand human instructions to manipulate everyday objects. Recent approaches in LGRM rely on pre-trained Visual Grounding (VG…

Lifelong learningObject DetectionVisual Grounding

Phoenix: A Motion-based Self-Reflection Framework for Fine-grained Robotic Action Correction

2025-04-20 · CVPR 2025 1 · Wenke Xia, Ruoxuan Feng, Dong Wang, Di Hu

Building a generalizable self-correction system is crucial for robots to recover from failures. Despite advancements in Multimodal Large Language Models (MLLMs) that empower robots with semantic reflection ability for fa…

Lifelong learning

Language-Conditioned Representations and Mixture-of-Experts Policy for Robust Multi-Task Robotic Manipulation

2025-10-28 · Xiucheng Zhang, Yang Jiang, Hongwei Qing, Jiashuo Bai arxiv

Perceptual ambiguity and task conflict limit multitask robotic manipulation via imitation learning. We propose a framework combining a Language-Conditioned Visual Representation (LCVR) module and a Language-conditioned M…