paper-with-me

Papers

OBJECT DYNAMICS DISTILLATION FOR SCENE DECOMPOSITION AND REPRESENTATION

2021-09-29 · ICLR 2022 4 · Qu Tang, Xiangyu Zhu, Zhen Lei, Zhaoxiang Zhang

The ability to perceive scenes in terms of abstract entities is crucial for us to achieve higher-level intelligence. Recently, several methods have been proposed to learn object-centric representations of scenes with multiple objects, yet most of which focus on static scenes. In this paper, we work on object dynamics and propose Object Dynamics Distillation Network (ODDN), a framework that distillates explicit object dynamics (e.g., velocity) from sequential static representations. ODDN also builds a relation module to model object interactions. We verify our approach on tasks of video reasoning and video prediction, which are two important evaluations for video understanding. The results show that the reasoning model with visual representations of ODDN performs better in answering reasoning questions around physical events in a video compared to the previous state-of-the-art methods. The distilled object dynamics also could be used to predict future video frames given two input frames, involving occlusion and objects collision. In addition, our architecture brings better segmentation quality and higher reconstruction accuracy.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectPredict Future Video FramesVideo PredictionVideo Understanding

Similar Papers 제목 키워드 기반

Unsupervised Video Decomposition using Spatio-temporal Iterative Inference

2020-06-25 · Polina Zablotskaia, Edoardo A. Dominici, Leonid Sigal, Andreas M. Lehrmann

Unsupervised multi-object scene decomposition is a fast-emerging problem in representation learning. Despite significant progress in static scenes, such models are unable to leverage important dynamic cues present in vid…

ObjectRepresentation Learning

Proactive Scene Decomposition and Reconstruction

2025-10-17 · Baicheng Li, Zike Yan, Dong Wu, Hongbin Zha arxiv

Human behaviors are the major causes of scene dynamics and inherently contain rich cues regarding the dynamics. This paper formalizes a new task of proactive scene decomposition and reconstruction, an online approach tha…

Pose Estimation

OBJECT-ORIENTED REPRESENTATION OF 3D SCENES

2019-09-25 · Chang Chen, Sungjin Ahn

In this paper, we propose a generative model, called ROOTS (Representation of Object-Oriented Three-dimension Scenes), for unsupervised object-wise 3D-scene decomposition and and rendering. For 3D scene modeling, ROOTS b…

DisentanglementObject

Learning Unified Decompositional and Compositional NeRF for Editable Novel View Synthesis

2023-08-05 · ICCV 2023 1 · Yuxin Wang, Wayne Wu, Dan Xu

Implicit neural representations have shown powerful capacity in modeling real-world 3D scenes, offering superior performance in novel view synthesis. In this paper, we target a more challenging scenario, i.e., joint scen…

NeRFNovel View Synthesis

${M^2D}$NeRF: Multi-Modal Decomposition NeRF with 3D Feature Fields

2024-05-08 · Ning Wang, Lefei Zhang, Angel X Chang

Neural fields (NeRF) have emerged as a promising approach for representing continuous 3D scenes. Nevertheless, the lack of semantic encoding in NeRFs poses a significant challenge for scene decomposition. To address this…

NeRF