paper-with-me

홈 › Papers

Robot Learning from Any Images

2025-09-26 · Siheng Zhao, Jiageng Mao, Wei Chow, Zeyu Shangguan, Tianheng Shi, Rong Xue, Yuxi Zheng, Yijia Weng, Yang You, Daniel Seita, Leonidas Guibas, Sergey Zakharov, Vitor Guizilini, Yue Wang arxiv

We introduce RoLA, a framework that transforms any in-the-wild image into an interactive, physics-enabled robotic environment. Unlike previous methods, RoLA operates directly on a single image without requiring additional hardware or digital assets. Our framework democratizes robotic data generation by producing massive visuomotor robotic demonstrations within minutes from a wide range of image sources, including camera captures, robotic datasets, and Internet images. At its core, our approach combines a novel method for single-view physical scene recovery with an efficient visual blending strategy for photorealistic data collection. We demonstrate RoLA's versatility across applications like scalable robotic data generation and augmentation, robot learning from Internet images, and single-image real-to-sim-to-real systems for manipulators and humanoids. Video results are available at https://sihengz02.github.io/RoLA .

📄 PDF Abstract BibTeX arXiv:2509.22970

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Research and Design on Intelligent Recognition of Unordered Targets for Robots Based on Reinforcement Learning

2025-03-10 · Yiting Mao, Dajun Tao, Shengyuan Zhang, Tian Qi 외

In the field of robot target recognition research driven by artificial intelligence (AI), factors such as the disordered distribution of targets, the complexity of the environment, the massive scale of data, and noise in…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning

Machine Learning in Appearance-based Robot Self-localization

2017-06-17 · Alexander Kuleshov, Alexander Bernstein, Evgeny Burnaev, Yury Yanovich

An appearance-based robot self-localization problem is considered in the machine learning framework. The appearance space is composed of all possible images, which can be captured by a robot's visual system under all rob…

BIG-bench Machine Learning

Real-time Pose Estimation from Images for Multiple Humanoid Robots

2021-07-06 · Arash Amini, Hafez Farazi, Sven Behnke

Pose estimation commonly refers to computer vision methods that recognize people's body postures in images or videos. With recent advancements in deep learning, we now have compelling models to tackle the problem in real…

Pose Estimation

Visual Generalized Coordinates

2015-09-18 · M. Seetha Ramaiah, Amitabha Mukerjee, Arindam Chakraborty, Sadbodh Sharma

An open problem in robotics is that of using vision to identify a robot's own body and the world around it. Many models attempt to recover the traditional C-space parameters. Instead, we propose an alternative C-space by…

3D Robot Pose Estimation from 2D Images

2019-02-13 · Christoph Heindl, Sebastian Zambal, Thomas Ponitz, Andreas Pichler 외

This paper considers the task of locating articulated poses of multiple robots in images. Our approach simultaneously infers the number of robots in a scene, identifies joint locations and estimates sparse depth maps aro…

Pose EstimationRobot Pose Estimation