paper-with-me

홈 › Papers

Time-Aware and View-Aware Video Rendering for Unsupervised Representation Learning

2018-11-26 · Shruti Vyas, Yogesh S Rawat, Mubarak Shah

The recent success in deep learning has lead to various effective representation learning methods for videos. However, the current approaches for video representation require large amount of human labeled datasets for effective learning. We present an unsupervised representation learning framework to encode scene dynamics in videos captured from multiple viewpoints. The proposed framework has two main components: Representation Learning Network (RL-NET), which learns a representation with the help of Blending Network (BL-NET), and Video Rendering Network (VR-NET), which is used for video synthesis. The framework takes as input video clips from different viewpoints and time, learns an internal representation and uses this representation to render a video clip from an arbitrary given viewpoint and time. The ability of the proposed network to render video frames from arbitrary viewpoints and time enable it to learn a meaningful and robust representation of the scene dynamics. We demonstrate the effectiveness of the proposed method in rendering view-aware as well as time-aware video clips on two different real-world datasets including UCF-101 and NTU-RGB+D. To further validate the effectiveness of the learned representation, we use it for the task of view-invariant activity classification where we observe a significant improvement (~26%) in the performance on NTU-RGB+D dataset compared to the existing state-of-the art methods.

📄 PDF Abstract BibTeX arXiv:1811.10699

Code (0)

등록된 구현이 없습니다.

Tasks

Representation Learning

Similar Papers 제목 키워드 기반

iButter: Neural Interactive Bullet Time Generator for Human Free-viewpoint Rendering

2021-08-12 · Liao Wang, Ziyu Wang, Pei Lin, Yuheng Jiang 외

Generating ``bullet-time'' effects of human free-viewpoint videos is critical for immersive visual effects and VR/AR experience. Recent neural advances still lack the controllable and interactive bullet-time design abili…

NeRFVideo Generation

Ground4D: Consistency-Aware 4D Reconstruction from Monocular Video

2026-06-27 · Qing Zhao, Weijian Deng, Pengxu Wei, Liang Lin arxiv

Learning a 4D scene representation from a single monocular video that supports dynamic novel-view synthesis while maintaining faithful geometry over time remains challenging. Dynamic Gaussian Splatting achieves strong re…

3D-Aware Video Generation

2022-06-29 · Sherwin Bahmani, Jeong Joon Park, Despoina Paschalidou, Hao Tang 외

Generative models have emerged as an essential building block for many image synthesis and editing tasks. Recent advances in this field have also enabled high-quality 3D or video content to be generated that exhibits eit…

Image GenerationVideo Generation

3DTV: A Feedforward Interpolation Network for Real-Time View Synthesis

2026-04-13 · Stefan Schulz, Fernando Edelstein, Hannah Dröge, Matthias B. Hullin 외 arxiv

Real-time free-viewpoint rendering requires balancing multi-camera redundancy with the latency constraints of interactive applications. We address this challenge by combining lightweight geometry with learning and propos…

MoVerse: Real-Time Video World Modeling with Panoramic Gaussian Scaffold

2026-06-11 · Yang Zhou, Ziheng Wang, Yuqin Lu, Haofeng Liu 외 arxiv

We present MoVerse, a real-time video world model that creates an interactively navigable scene from a single narrow-field-of-view image. This setting is challenging because the input observes only a small fraction of th…