paper-with-me

홈 › Papers

Robust Camera Pose Refinement for Multi-Resolution Hash Encoding

2023-02-03 · Hwan Heo, Taekyung Kim, Jiyoung Lee, Jaewon Lee, Soohyun Kim, Hyunwoo J. Kim, Jin-Hwa Kim

Multi-resolution hash encoding has recently been proposed to reduce the computational cost of neural renderings, such as NeRF. This method requires accurate camera poses for the neural renderings of given scenes. However, contrary to previous methods jointly optimizing camera poses and 3D scenes, the naive gradient-based camera pose refinement method using multi-resolution hash encoding severely deteriorates performance. We propose a joint optimization algorithm to calibrate the camera pose and learn a geometric representation using efficient multi-resolution hash encoding. Showing that the oscillating gradient flows of hash encoding interfere with the registration of camera poses, our method addresses the issue by utilizing smooth interpolation weighting to stabilize the gradient oscillation for the ray samplings across hash grids. Moreover, the curriculum training procedure helps to learn the level-wise hash encoding, further increasing the pose refinement. Experiments on the novel-view synthesis datasets validate that our learning frameworks achieve state-of-the-art performance and rapid convergence of neural rendering, even when initial camera poses are unknown.

📄 PDF Abstract BibTeX arXiv:2302.01571

Code (0)

등록된 구현이 없습니다.

Tasks

NeRFNeural RenderingNovel View Synthesis

Similar Papers 제목 키워드 기반

Model Inversion Attack Against Deep Hashing

2025-11-15 · Dongdong Zhao, Qiben Xu, Ranxin Fang, Baogang Song arxiv

Deep hashing improves retrieval efficiency through compact binary codes, yet it introduces severe and often overlooked privacy risks. The ability to reconstruct original training data from hash codes could lead to seriou…

Mesh-based Camera Pairs Selection and Occlusion-Aware Masking for Mesh Refinement

2019-05-21 · Andrea Romanoni, Matteo Matteucci

Many Multi-View-Stereo algorithms extract a 3D mesh model of a scene, after fusing depth maps into a volumetric representation of the space. Due to the limited scalability of such representations, the estimated model doe…

NeRSemble: Multi-view Radiance Field Reconstruction of Human Heads

2023-05-04 · Tobias Kirschstein, Shenhan Qian, Simon Giebenhain, Tim Walter 외

We focus on reconstructing high-fidelity radiance fields of human heads, capturing their animations over time, and synthesizing re-renderings from novel viewpoints at arbitrary time steps. To this end, we propose a new m…

NeuV-SLAM: Fast Neural Multiresolution Voxel Optimization for RGBD Dense SLAM

2024-02-03 · Wenzhi Guo, Bing Wang, Lijun Chen

We introduce NeuV-SLAM, a novel dense simultaneous localization and mapping pipeline based on neural multiresolution voxels, characterized by ultra-fast convergence and incremental expansion capabilities. This pipeline u…

Simultaneous Localization and Mapping

Spk2SRImgNet: Super-Resolve Dynamic Scene from Spike Stream via Motion Aligned Collaborative Filtering

2025-01-01 · CVPR 2025 1 · Yuanlin Wang, Yiyang Zhang, Ruiqin Xiong, Jing Zhao 외

Spike camera is a kind of neuromorphic camera that records dynamic scenes by firing a stream of binary spikes with extremely high temporal resolution. It demonstrates great potential for vision tasks in high-speed sc…

Collaborative FilteringImage RestorationSuper-Resolution