paper-with-me

홈 › Papers

Predicting 3D representations for Dynamic Scenes

2025-01-28 · Di Qi, Tong Yang, Beining Wang, Xiangyu Zhang, Wenqiang Zhang

We present a novel framework for dynamic radiance field prediction given monocular video streams. Unlike previous methods that primarily focus on predicting future frames, our method goes a step further by generating explicit 3D representations of the dynamic scene. The framework builds on two core designs. First, we adopt an ego-centric unbounded triplane to explicitly represent the dynamic physical world. Second, we develop a 4D-aware transformer to aggregate features from monocular videos to update the triplane. Coupling these two designs enables us to train the proposed model with large-scale monocular videos in a self-supervised manner. Our model achieves top results in dynamic radiance field prediction on NVIDIA dynamic scenes, demonstrating its strong performance on 4D physical world modeling. Besides, our model shows a superior generalizability to unseen scenarios. Notably, we find that our approach emerges capabilities for geometry and semantic learning.

📄 PDF Abstract BibTeX arXiv:2501.16617

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

ADOPT Please enter a description about the method here
Focus 설명 없음

Similar Papers 제목 키워드 기반

Neural Scene Graphs for Dynamic Scenes

2020-11-20 · CVPR 2021 1 · Julian Ost, Fahim Mannan, Nils Thuerey, Julian Knodt 외

Recent implicit neural rendering methods have demonstrated that it is possible to learn accurate view synthesis for complex scenes by predicting their volumetric density and color supervised solely by a set of RGB images…

Neural Rendering

Representing Volumetric Videos as Dynamic MLP Maps

2023-04-13 · CVPR 2023 1 · Sida Peng, Yunzhi Yan, Qing Shuai, Hujun Bao 외

This paper introduces a novel representation of volumetric videos for real-time view synthesis of dynamic scenes. Recent advances in neural scene representations demonstrate their remarkable capability to model and rende…

DecoderGPU

Dynamic Scene Understanding from Vision-Language Representations

2025-01-20 · Shahaf Pruss, Morris Alper, Hadar Averbuch-Elor

Images depicting complex, dynamic scenes are challenging to parse automatically, requiring both high-level comprehension of the overall situation and fine-grained identification of participating entities and their intera…

Grounded Situation RecognitionHuman-Human Interaction RecognitionHuman Interaction RecognitionHuman-Object Interaction Detection+2

SlotGNN: Unsupervised Discovery of Multi-Object Representations and Visual Dynamics

2023-10-06 · Alireza Rezazadeh, Athreyi Badithela, Karthik Desingh, Changhyun Choi

Learning multi-object dynamics from visual data using unsupervised techniques is challenging due to the need for robust, object representations that can be learned through robot interactions. This paper presents a novel …

ObjectObject DiscoveryObject RearrangementSpatial Reasoning

Trajectory Forecasting on Temporal Graphs

2022-07-01 · Görkay Aydemir, Adil Kaan Akan, Fatma Güney

Predicting future locations of agents in the scene is an important problem in self-driving. In recent years, there has been a significant progress in representing the scene and the agents in it. The interactions of agent…

Graph Neural NetworkMotion ForecastingTrajectory Forecasting