paper-with-me

Papers

Coupled Distributional Random Expert Distillation for World Model Online Imitation Learning

2025-05-04 · Shangzhe Li, Zhiao Huang, Hao Su

Imitation Learning (IL) has achieved remarkable success across various domains, including robotics, autonomous driving, and healthcare, by enabling agents to learn complex behaviors from expert demonstrations. However, existing IL methods often face instability challenges, particularly when relying on adversarial reward or value formulations in world model frameworks. In this work, we propose a novel approach to online imitation learning that addresses these limitations through a reward model based on random network distillation (RND) for density estimation. Our reward model is built on the joint estimation of expert and behavioral distributions within the latent space of the world model. We evaluate our method across diverse benchmarks, including DMControl, Meta-World, and ManiSkill2, showcasing its ability to deliver stable performance and achieve expert-level results in both locomotion and manipulation tasks. Our approach demonstrates improved stability over adversarial methods while maintaining expert-level performance.

📄 PDF Abstract BibTeX arXiv:2505.02228

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDensity EstimationImitation Learning

Similar Papers 제목 키워드 기반

Distillation-based Scenario-Adaptive Mixture-of-Experts for the Matching Stage of Multi-scenario Recommendation

2025-11-28 · Ruibing Wang, Shuhan Guo, Haotong Du, Quanming Yao arxiv

Multi-scenario recommendation is pivotal for optimizing user experience across diverse contexts. While Multi-gate Mixture-of-Experts (MMOE) thrives in ranking, its transfer to the matching stage is hindered by the blind …

Knowledge Distillation

Distilling Long-tailed Datasets

2024-08-24 · CVPR 2025 1 · Zhenghao Zhao, Haoxuan Wang, Yuzhang Shang, Kai Wang 외

Dataset distillation (DD) aims to distill a small, information-rich dataset from a larger one for efficient neural network training. However, existing DD methods struggle with long-tailed datasets, which are prevalent in…

Dataset DistillationEfficient Neural Network

Decoupled Multimodal Distilling for Emotion Recognition

2023-03-24 · CVPR 2023 1 · Yong Li, Yuanzhi Wang, Zhen Cui

Human multimodal emotion recognition (MER) aims to perceive human emotions via language, visual and acoustic modalities. Despite the impressive performance of previous MER approaches, the inherent multimodal heterogeneit…

Emotion RecognitionKnowledge DistillationMultimodal Emotion RecognitionTransfer Learning

Distributional Dataset Distillation with Subtask Decomposition

2024-03-01 · Tian Qin, Zhiwei Deng, David Alvarez-Melis

What does a neural network learn when training from a task-specific dataset? Synthesizing this knowledge is the central idea behind Dataset Distillation, which recent work has shown can be used to compress large datasets…

Dataset DistillationDecoder

Uncertainty-Aware Multi-Expert Knowledge Distillation for Imbalanced Disease Grading

2025-05-01 · Shuo Tong, Shangde Gao, Ke Liu, Zihang Huang 외

Automatic disease image grading is a significant application of artificial intelligence for healthcare, enabling faster and more accurate patient assessments. However, domain shifts, which are exacerbated by data imbalan…

Knowledge DistillationTransfer Learning