paper-with-me

Papers

PlayWorld: Learning Robot World Models from Autonomous Play

2026-03-09 · Tenny Yin, Zhiting Mei, Zhonghe Zheng, Miyu Yamane, David Wang, Jade Sceats, Samuel M. Bateman, Lihan Zha, Apurva Badithela, Ola Shorinwa, Anirudha Majumdar arxiv

Action-conditioned video models offer a promising path to building general-purpose robot simulators that can improve directly from data. Yet, despite training on large-scale robot datasets, current state-of-the-art video models still struggle to predict physically consistent robot-object interactions that are crucial in robotic manipulation. To close this gap, we present PlayWorld, a simple, scalable, and fully autonomous pipeline for training high-fidelity video world simulators from interaction experience. In contrast to prior approaches that rely on success-biased human demonstrations, PlayWorld is the first system capable of learning entirely from unsupervised robot self-play, enabling naturally scalable data collection while capturing complex, long-tailed physical interactions essential for modeling realistic object dynamics. Experiments across diverse manipulation tasks show that PlayWorld generates high-quality, physically consistent predictions for contact-rich interactions that are not captured by world models trained on human-collected data. We further demonstrate the versatility of PlayWorld in enabling fine-grained failure prediction and policy evaluation, with up to 40% improvements over human-collected data. Finally, we demonstrate how PlayWorld enables reinforcement learning in the world model, improving policy performance by 65% in success rates when deployed in the real world.

📄 PDF Abstract BibTeX arXiv:2603.09030

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives

2026-08-13 · Kaixin Ding, Xi Chen, Minghong Cai, Zhiyuan Xu 외 arxiv

Video world models simulate future states conditioned on current observations and user actions. Recent systems have demonstrated impressive video consistency and action controllability over long sequences. However, fairl…

ALAN: Autonomously Exploring Robotic Agents in the Real World

2023-02-13 · Russell Mendonca, Shikhar Bahl, Deepak Pathak

Robotic agents that operate autonomously in the real world need to continuously explore their environment and learn from the data collected, with minimal human supervision. While it is possible to build agents that can l…

Tether: Autonomous Functional Play with Correspondence-Driven Trajectory Warping

2026-03-03 · William Liang, Sam Wang, Hung-Ju Wang, Osbert Bastani 외 arxiv

The ability to conduct and learn from interaction and experience is a central challenge in robotics, offering a scalable alternative to labor-intensive human demonstrations. However, realizing such "play" requires (1) a …

Artificial Intelligence for Long-Term Robot Autonomy: A Survey

2018-07-13 · Lars Kunze, Nick Hawes, Tom Duckett, Marc Hanheide 외

Autonomous systems will play an essential role in many applications across diverse domains including space, marine, air, field, road, and service robotics. They will assist us in our daily routines and perform dangerous,…

Survey

A Data-Efficient Deep Learning Approach for Deployable Multimodal Social Robots

2019-08-27 · Heriberto Cuayáhuitl

The deep supervised and reinforcement learning paradigms (among others) have the potential to endow interactive multimodal social robots with the ability of acquiring skills autonomously. But it is still not very clear y…

Reinforcement LearningRobot Manipulation