paper-with-me

홈 › Papers

Diverse Skill Discovery for Quadruped Robots via Unsupervised Learning

2026-02-10 · Ruopeng Cui, Yifei Bi, Haojie Luo, Wei Li arxiv

Reinforcement learning necessitates meticulous reward shaping by specialists to elicit target behaviors, while imitation learning relies on costly task-specific data. In contrast, unsupervised skill discovery can potentially reduce these burdens by learning a diverse repertoire of useful skills driven by intrinsic motivation. However, existing methods exhibit two key limitations: they typically rely on a single policy to master a versatile repertoire of behaviors without modeling the shared structure or distinctions among them, which results in low learning efficiency; moreover, they are susceptible to reward hacking, where the reward signal increases and converges rapidly while the learned skills display insufficient actual diversity. In this work, we introduce an Orthogonal Mixture-of-Experts (OMoE) architecture that prevents diverse behaviors from collapsing into overlapping representations, enabling a single policy to master a wide spectrum of locomotion skills. In addition, we design a multi-discriminator framework in which different discriminators operate on distinct observation spaces, effectively mitigating reward hacking. We evaluated our method on the 12-DOF Unitree A1 quadruped robot, demonstrating a diverse set of locomotion skills. Our experiments demonstrate that the proposed framework boosts training efficiency and yields an 18.3\% expansion in state-space coverage compared to the baseline.

📄 PDF Abstract BibTeX arXiv:2602.09767

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion

2025-08-12 · Seungeun Rho, Kartik Garg, Morgan Byrd, Sehoon Ha arxiv

Exploration is crucial for enabling legged robots to learn agile locomotion behaviors that can overcome diverse obstacles. However, such exploration is inherently challenging, and we often rely on extensive reward engine…

From Tabula Rasa to Emergent Abilities: Discovering Robot Skills via Real-World Unsupervised Quality-Diversity

2025-08-26 · Luca Grillotti, Lisa Coiffard, Oscar Pang, Maxence Faldor 외 arxiv

Autonomous skill discovery aims to enable robots to acquire diverse behaviors without explicit supervision. Learning such behaviors directly on physical hardware remains challenging due to safety and data efficiency cons…

Versatile Skill Control via Self-supervised Adversarial Imitation of Unlabeled Mixed Motions

2022-09-16 · Chenhao Li, Sebastian Blaes, Pavel Kolev, Marin Vlastelica 외

Learning diverse skills is one of the main challenges in robotics. To this end, imitation learning approaches have achieved impressive results. These methods require explicitly labeled datasets or assume consistent skill…

Imitation Learning

KiRAS: Keyframe Guided Self-Imitation for Robust and Adaptive Skill Learning in Quadruped Robots

2026-03-16 · Xiaoyi Wei, Peng Zhai, Jiaxin Tu, Yueqi Zhang 외 arxiv

With advances in reinforcement learning and imitation learning, quadruped robots can acquire diverse skills within a single policy by imitating multiple skill-specific datasets. However, the lack of datasets on complex t…

Reinforcement Learning

GenLoco: Generalized Locomotion Controllers for Quadrupedal Robots

2022-09-12 · Gilbert Feng, Hongbo Zhang, Zhongyu Li, Xue Bin Peng 외

Recent years have seen a surge in commercially-available and affordable quadrupedal robots, with many of these platforms being actively used in research and industry. As the availability of legged robots grows, so does t…