paper-with-me

홈 › Papers

3D-OES: Viewpoint-Invariant Object-Factorized Environment Simulators

2020-11-12 · Hsiao-Yu Fish Tung, Zhou Xian, Mihir Prabhudesai, Shamit Lal, Katerina Fragkiadaki

We propose an action-conditioned dynamics model that predicts scene changes caused by object and agent interactions in a viewpoint-invariant 3D neural scene representation space, inferred from RGB-D videos. In this 3D feature space, objects do not interfere with one another and their appearance persists over time and across viewpoints. This permits our model to predict future scenes long in the future by simply "moving" 3D object features based on cumulative object motion predictions. Object motion predictions are computed by a graph neural network that operates over the object features extracted from the 3D neural scene representation. Our model's simulations can be decoded by a neural renderer into2D image views from any desired viewpoint, which aids the interpretability of our latent 3D simulation space. We show our model generalizes well its predictions across varying number and appearances of interacting objects as well as across camera viewpoints, outperforming existing 2D and 3D dynamics models. We further demonstrate sim-to-real transfer of the learnt dynamics by applying our model trained solely in simulation to model-based control for pushing objects to desired locations under clutter on a real robotic setup

📄 PDF Abstract BibTeX arXiv:2011.06464

Code (0)

등록된 구현이 없습니다.

Tasks

Graph Neural NetworkObject

Methods 이 논문이 사용한 방법론

Graph Neural Network 설명 없음
Interpretability 설명 없음

Similar Papers 제목 키워드 기반

3D-GIF: 3D-Controllable Object Generation via Implicit Factorized Representations

2022-03-12 · Minsoo Lee, Chaeyeon Chung, Hojun Cho, Minjung Kim 외

While NeRF-based 3D-aware image generation methods enable viewpoint control, limitations still remain to be adopted to various 3D applications. Due to their view-dependent and light-entangled volume representation, the 3…

3D geometryImage GenerationNeRF

ConTraIRL: Factorized Contrastive Abstractions for Transferable IRL

2026-06-02 · Yikang Gui, Bikramjit Banerjee, Prashant Doshi arxiv

Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and task goals. We propose Factorized Contrastive Abstractions for Transf…

Reinforcement LearningContinuous Control

Graph-Structured Visual Imitation

2019-07-11 · Maximilian Sieb, Zhou Xian, Audrey Huang, Oliver Kroemer 외

We cast visual imitation as a visual correspondence problem. Our robotic agent is rewarded when its actions result in better matching of relative spatial configurations for corresponding visual entities detected in its w…

Learning Fine-grained View-Invariant Representations from Unpaired Ego-Exo Videos via Temporal Alignment

2023-06-08 · NeurIPS 2023 11

The egocentric and exocentric viewpoints of a human activity look dramatically different, yet invariant representations to link them are essential for many potential applications in robotics and augmented reality. Prior …

Video Understanding

View-Invariant Localization using Semantic Objects in Changing Environments

2022-09-28 · Jacqueline Ankenbauer, Kaveh Fathian, Jonathan P. How

This paper proposes a novel framework for real-time localization and egomotion tracking of a vehicle in a reference map. The core idea is to map the semantic objects observed by the vehicle and register them to their cor…

Position