paper-with-me

홈 › Papers

Unveiling the Impact of Data and Model Scaling on High-Level Control for Humanoid Robots

2025-11-12 · Yuxi Wei, Zirui Wang, Kangning Yin, Yue Hu, Jingbo Wang, Siheng Chen arxiv

Data scaling has long remained a critical bottleneck in robot learning. For humanoid robots, human videos and motion data are abundant and widely available, offering a free and large-scale data source. Besides, the semantics related to the motions enable modality alignment and high-level robot control learning. However, how to effectively mine raw video, extract robot-learnable representations, and leverage them for scalable learning remains an open problem. To address this, we introduce Humanoid-Union, a large-scale dataset generated through an autonomous pipeline, comprising over 260 hours of diverse, high-quality humanoid robot motion data with semantic annotations derived from human motion videos. The dataset can be further expanded via the same pipeline. Building on this data resource, we propose SCHUR, a scalable learning framework designed to explore the impact of large-scale data on high-level control in humanoid robots. Experimental results demonstrate that SCHUR achieves high robot motion generation quality and strong text-motion alignment under data and model scaling, with 37\% reconstruction improvement under MPJPE and 25\% alignment improvement under FID comparing with previous methods. Its effectiveness is further validated through deployment in real-world humanoid robot.

📄 PDF Abstract BibTeX arXiv:2511.09241

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unveiling Chain of Step Reasoning for Vision-Language Models with Fine-grained Rewards

2025-09-23 · Honghao Chen, Xingzhou Lou, Xiaokun Feng, Kaiqi Huang 외 arxiv

Chain of thought reasoning has demonstrated remarkable success in large language models, yet its adaptation to vision-language reasoning remains an open challenge with unclear best practices. Existing attempts typically …

Reinforcement LearningMultimodal Reasoning

Unveiling Scaling Behaviors in Molecular Language Models: Effects of Model Size, Data, and Representation

2026-01-30 · Dong Xu, Qihua Pan, Sisi Yuan, Jianqiang Li 외 arxiv

Molecular generative models, often employing GPT-style language modeling on molecular string representations, have shown promising capabilities when scaled to large datasets and model sizes. However, it remains unclear a…

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments

2026-07-06 · Deyao Zhu, Xin Zhou, Shengling Qin, Xuekai Zhu 외 arxiv

Pretraining scaling laws reveal that model capability improves predictably with data and compute. But learning from real world environments after deployment remains far less understood. Analyzing roughly 38,000 hours of …

Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought

2025-09-04 · Zeyu Gan, Hao Yi, Yong Liu arxiv

Test-time scaling, primarily manifested through multi-step Chain-of-Thought (CoT) reasoning via Reinforcement Learning (RL), has emerged as a pivotal paradigm for enhancing the reasoning capabilities of Large Language Mo…

Reinforcement Learning

FreSca: Unveiling the Scaling Space in Diffusion Models

2025-04-02 · Chao Huang, Susan Liang, Yunlong Tang, Li Ma 외

Diffusion models offer impressive controllability for image tasks, primarily through noise predictions that encode task-specific information and classifier-free guidance enabling adjustable scaling. This scaling mechanis…

Depth Estimation