paper-with-me

홈 › Papers

PS4PRO: Pixel-to-pixel Supervision for Photorealistic Rendering and Optimization

2025-05-28 · Yezhi Shen, Qiuchen Zhai, Fengqing Zhu

Neural rendering methods have gained significant attention for their ability to reconstruct 3D scenes from 2D images. The core idea is to take multiple views as input and optimize the reconstructed scene by minimizing the uncertainty in geometry and appearance across the views. However, the reconstruction quality is limited by the number of input views. This limitation is further pronounced in complex and dynamic scenes, where certain angles of objects are never seen. In this paper, we propose to use video frame interpolation as the data augmentation method for neural rendering. Furthermore, we design a lightweight yet high-quality video frame interpolation model, PS4PRO (Pixel-to-pixel Supervision for Photorealistic Rendering and Optimization). PS4PRO is trained on diverse video datasets, implicitly modeling camera movement as well as real-world 3D geometry. Our model performs as an implicit world prior, enriching the photo supervision for 3D reconstruction. By leveraging the proposed method, we effectively augment existing datasets for neural rendering methods. Our experimental results indicate that our method improves the reconstruction performance on both static and dynamic scenes.

📄 PDF Abstract BibTeX arXiv:2505.22616

Code (0)

등록된 구현이 없습니다.

Tasks

3D geometry3D ReconstructionData AugmentationNeural RenderingVideo Frame Interpolation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Gaussian Pixel Codec Avatars: A Hybrid Representation for Efficient Rendering

2025-12-17 · Divam Gupta, Anuj Pahuja, Nemanja Bartolovic, Tomas Simon 외 arxiv

We present Gaussian Pixel Codec Avatars (GPiCA), photorealistic head avatars that can be generated from multi-view images and efficiently rendered on mobile devices. GPiCA utilizes a unique hybrid representation that com…

Voxify3D: Pixel Art Meets Volumetric Rendering

2025-12-08 · Yi-Chuan Huang, Jiewen Chan, Hao-Jen Chien, Yu-Lun Liu arxiv

Voxel art is a distinctive stylization widely used in games and digital media, yet automated generation from 3D meshes remains challenging due to conflicting requirements of geometric abstraction, semantic preservation, …

${C}^{3}$-GS: Learning Context-aware, Cross-dimension, Cross-scale Feature for Generalizable Gaussian Splatting

2025-08-28 · Yuxi Hu, Jun Zhang, Kuangyi Chen, Zhe Zhang 외 arxiv

Generalizable Gaussian Splatting aims to synthesize novel views for unseen scenes without per-scene optimization. In particular, recent advancements utilize feed-forward networks to predict per-pixel Gaussian parameters,…

TurboGS: Accelerating 3D Gaussian Splatting via Error-Guided Sparse Pixel Sampling and Optimization

2026-06-14 · Zheng Dong, Daifei Qiu, Pinxuan Dai, Ke Xu 외 arxiv

Consumer-level applications require fast optimization of 3D Gaussian Splatting (3DGS) with high-fidelity novel view rendering. However, existing 3DGS acceleration approaches still incur substantial computation on redunda…

StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models

2024-12-17 · CVPR 2025 1 · Yunzhi Yan, Zhen Xu, Haotong Lin, Haian Jin 외

This paper aims to tackle the problem of photorealistic view synthesis from vehicle sensor data. Recent advancements in neural scene representation have achieved notable success in rendering high-quality autonomous drivi…

Autonomous DrivingNovel View Synthesis