paper-with-me

Papers

T-Code: Simple Temporal Latent Code for Efficient Dynamic View Synthesis

2023-12-18 · Zhenhuan Liu, Shuai Liu, Jie Yang, Wei Liu

Novel view synthesis for dynamic scenes is one of the spotlights in computer vision. The key to efficient dynamic view synthesis is to find a compact representation to store the information across time. Though existing methods achieve fast dynamic view synthesis by tensor decomposition or hash grid feature concatenation, their mixed representations ignore the structural difference between time domain and spatial domain, resulting in sub-optimal computation and storage cost. This paper presents T-Code, the efficient decoupled latent code for the time dimension only. The decomposed feature design enables customizing modules to cater for different scenarios with individual specialty and yielding desired results at lower cost. Based on T-Code, we propose our highly compact hybrid neural graphics primitives (HybridNGP) for multi-camera setting and deformation neural graphics primitives with T-Code (DNGP-T) for monocular scenario. Experiments show that HybridNGP delivers high fidelity results at top processing speed with much less storage consumption, while DNGP-T achieves state-of-the-art quality and high training speed for monocular reconstruction.

📄 PDF Abstract BibTeX arXiv:2312.11015

Code (0)

등록된 구현이 없습니다.

Tasks

Monocular ReconstructionNovel View SynthesisTensor Decomposition

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Sparse identification of nonlinear dynamics and Koopman operators with Shallow Recurrent Decoder Networks

2025-01-23 · Mars Liyao Gao, Jan P. Williams, J. Nathan Kutz

Modeling real-world spatio-temporal data is exceptionally difficult due to inherent high dimensionality, measurement noise, partial observations, and often expensive data collection procedures. In this paper, we present …

Decoder

Speech Modeling with a Hierarchical Transformer Dynamical VAE

2023-03-07 · Xiaoyu Lin, Xiaoyu Bie, Simon Leglaive, Laurent Girin 외

The dynamical variational autoencoders (DVAEs) are a family of latent-variable deep generative models that extends the VAE to model a sequence of observed data and a corresponding sequence of latent vectors. In almost al…

Speech Enhancement

Adaptive 1D Video Diffusion Autoencoder

2026-02-04 · Yao Teng, Minxuan Lin, Xian Liu, Shuai Wang 외 arxiv

Recent video generation models largely rely on video autoencoders that compress pixel-space videos into latent representations. However, existing video autoencoders suffer from three major limitations: (1) fixed-rate com…

Video Generation

Video Generation with Predictive Latents

2026-05-04 · Yian Zhao, Feng Wang, Qiushan Guo, Chang Liu 외 arxiv

Video Variational Autoencoder (VAE) enables latent video generative modeling by mapping the visual world into compact spatiotemporal latent spaces, improving training efficiency and stability. While existing video VAEs a…

Video ReconstructionVideo Generation

Context-Aware Markov VAE for CSI Compression in Wireless Systems

2026-06-15 · Efstathios Chatziloizos, Konstantinos Vandikas, Aneta Vulgarakis Feljan, Zheng Chen 외 arxiv

This paper considers neural channel state information (CSI) compression for time-varying massive multiple-input multiple-output (MIMO) channels in frequency division duplex (FDD) systems with limited feedback resources. …