paper-with-me

Papers

Learning in ImaginationLand: Omnidirectional Policies through 3D Generative Models (OP-Gen)

2025-09-07 · Yifei Ren, Edward Johns arxiv

Recent 3D generative models, which are capable of generating full object shapes from just a few images, now open up new opportunities in robotics. In this work, we show that 3D generative models can be used to augment a dataset from a single real-world demonstration, after which an omnidirectional policy can be learned within this imagined dataset. We found that this enables a robot to perform a task when initialised from states very far from those observed during the demonstration, including starting from the opposite side of the object relative to the real-world demonstration, significantly reducing the number of demonstrations required for policy learning. Through several real-world experiments across tasks such as grasping objects, opening a drawer, and placing trash into a bin, we study these omnidirectional policies by investigating the effect of various design choices on policy behaviour, and we show superior performance to recent baselines which use alternative methods for data augmentation.

📄 PDF Abstract BibTeX arXiv:2509.06191

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentation

Similar Papers 제목 키워드 기반

Stay Seated: Learning Omnidirectional Humanoid Locomotion on a Passive Mobile Chair with Casters

2026-08-28 · Kango Yanagida, Kazuki Miyazawa, Takato Horii arxiv

Humanoid robots with quasi-direct-drive actuators continuously generate joint torque while standing, whereas seated humans delegate weight support to chairs during desk work. As a first step toward seated loco-manipulati…

Super-resolution of Omnidirectional Images Using Adversarial Learning

2019-08-12 · Cagri Ozcinar, Aakanksha Rana, Aljosa Smolic

An omnidirectional image (ODI) enables viewers to look in every direction from a fixed point through a head-mounted display providing an immersive experience compared to that of a standard image. Designing immersive virt…

Generative Adversarial NetworkSuper-Resolution

Bridge the Gap Between VQA and Human Behavior on Omnidirectional Video: A Large-Scale Dataset and a Deep Learning Model

2018-07-29 · Chen Li, Mai Xu, Xinzhe Du, Zulin Wang

Omnidirectional video enables spherical stimuli with the $360 \times 180^ \circ$ viewing range. Meanwhile, only the viewport region of omnidirectional video can be seen by the observer through head movement (HM), and an …

Visual Question Answering (VQA)

Gait in Eight: Efficient On-Robot Learning for Omnidirectional Quadruped Locomotion

2025-03-11 · Nico Bohlinger, Jonathan Kinzel, Daniel Palenicek, Lukasz Antczak 외

On-robot Reinforcement Learning is a promising approach to train embodiment-aware policies for legged robots. However, the computational constraints of real-time learning on robots pose a significant challenge. We presen…

Direct Triangulation with Spherical Projection for Omnidirectional Cameras

2022-06-08 · Ciarán Eising

In this paper, it is proposed to solve the problem of triangulation for calibrated omnidirectional cameras through the optimisation of ray-pairs on the projective sphere. The proposed solution boils down to finding the r…