paper-with-me

홈 › Papers

SHaDe: Compact and Consistent Dynamic 3D Reconstruction via Tri-Plane Deformation and Latent Diffusion

2025-05-22 · Asrar Alruwayqi

We present a novel framework for dynamic 3D scene reconstruction that integrates three key components: an explicit tri-plane deformation field, a view-conditioned canonical radiance field with spherical harmonics (SH) attention, and a temporally-aware latent diffusion prior. Our method encodes 4D scenes using three orthogonal 2D feature planes that evolve over time, enabling efficient and compact spatiotemporal representation. These features are explicitly warped into a canonical space via a deformation offset field, eliminating the need for MLP-based motion modeling. In canonical space, we replace traditional MLP decoders with a structured SH-based rendering head that synthesizes view-dependent color via attention over learned frequency bands improving both interpretability and rendering efficiency. To further enhance fidelity and temporal consistency, we introduce a transformer-guided latent diffusion module that refines the tri-plane and deformation features in a compressed latent space. This generative module denoises scene representations under ambiguous or out-of-distribution (OOD) motion, improving generalization. Our model is trained in two stages: the diffusion module is first pre-trained independently, and then fine-tuned jointly with the full pipeline using a combination of image reconstruction, diffusion denoising, and temporal consistency losses. We demonstrate state-of-the-art results on synthetic benchmarks, surpassing recent methods such as HexPlane and 4D Gaussian Splatting in visual quality, temporal coherence, and robustness to sparse-view dynamic inputs.

📄 PDF Abstract BibTeX arXiv:2505.16535

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction3D Scene ReconstructionDenoisingImage Reconstruction

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

WavePlanes: Compact Hex Planes for Dynamic Novel View Synthesis

2023-12-03 · Adrian Azzarelli, Nantheera Anantrasirichai, David R Bull

Dynamic Novel View Synthesis (Dynamic NVS) enhances NVS technologies to model moving 3-D scenes. However, current methods are resource intensive and challenging to compress. To address this, we present WavePlanes, a fast…

Novel View Synthesis

Pointmap Association and Piecewise-Plane Constraint for Consistent and Compact 3D Gaussian Segmentation Field

2025-02-22 · Wenhao Hu, Wenhao Chai, Shengyu Hao, Xiaotong Cui 외

Achieving a consistent and compact 3D segmentation field is crucial for maintaining semantic coherence across views and accurately representing scene structures. Previous 3D scene segmentation methods rely on video segme…

2D Panoptic Segmentation3D Scene ReconstructionPanoptic SegmentationScene Segmentation+3

Genetic Algorithms for Starshade Retargeting in Space-Based Telescopes

2019-07-23 · Ho Chit Siu, Victor Pankratius

Future space-based telescopes will leverage starshades as components that can be independently positioned. Starshades will adjust the light coming in from exoplanet host stars and enhance the direct imaging of exoplanets…

Scheduling

Neural LerPlane Representations for Fast 4D Reconstruction of Deformable Tissues

2023-05-31 · Chen Yang, Kailing Wang, Yuehao Wang, Xiaokang Yang 외

Reconstructing deformable tissues from endoscopic stereo videos in robotic surgery is crucial for various clinical applications. However, existing methods relying only on implicit representations are computationally expe…

4D reconstruction

Tensor4D: Efficient Neural 4D Decomposition for High-Fidelity Dynamic Reconstruction and Rendering

2023-01-01 · CVPR 2023 1 · Ruizhi Shao, Zerong Zheng, Hanzhang Tu, Boning Liu 외

We present Tensor4D, an efficient yet effective approach to dynamic scene modeling. The key of our solution is an efficient 4D tensor decomposition method so that the dynamic scene can be directly represented as a 4D…

Dynamic ReconstructionTensor Decomposition