paper-with-me

Papers

Nerfels: Renderable Neural Codes for Improved Camera Pose Estimation

2022-06-04 · Gil Avraham, Julian Straub, Tianwei Shen, Tsun-Yi Yang, Hugo Germain, Chris Sweeney, Vasileios Balntas, David Novotny, Daniel DeTone, Richard Newcombe

This paper presents a framework that combines traditional keypoint-based camera pose optimization with an invertible neural rendering mechanism. Our proposed 3D scene representation, Nerfels, is locally dense yet globally sparse. As opposed to existing invertible neural rendering systems which overfit a model to the entire scene, we adopt a feature-driven approach for representing scene-agnostic, local 3D patches with renderable codes. By modelling a scene only where local features are detected, our framework effectively generalizes to unseen local regions in the scene via an optimizable code conditioning mechanism in the neural renderer, all while maintaining the low memory footprint of a sparse 3D map representation. Our model can be incorporated to existing state-of-the-art hand-crafted and learned local feature pose estimators, yielding improved performance when evaluating on ScanNet for wide camera baseline scenarios.

📄 PDF Abstract BibTeX arXiv:2206.01916

Code (0)

등록된 구현이 없습니다.

Tasks

Camera Pose EstimationNeural RenderingPose Estimation

Similar Papers 제목 키워드 기반

Language-enhanced RNR-Map: Querying Renderable Neural Radiance Field maps with natural language

2023-08-17 · Francesco Taioli, Federico Cunico, Federico Girella, Riccardo Bologna 외

We present Le-RNR-Map, a Language-enhanced Renderable Neural Radiance map for Visual Navigation with natural language query prompts. The recently proposed RNR-Map employs a grid structure comprising latent codes position…

Language ModelingLanguage ModellingLarge Language ModelVisual Navigation

Renderable Neural Radiance Map for Visual Navigation

2023-03-01 · CVPR 2023 1 · Obin Kwon, Jeongho Park, Songhwai Oh

We propose a novel type of map for visual navigation, a renderable neural radiance map (RNR-Map), which is designed to contain the overall visual information of a 3D environment. The RNR-Map has a grid form and consists …

DescriptiveVisual LocalizationVisual Navigation

TryOnCrafter: Unleashing Camera Trajectories for Realistic Video Virtual Try-on via a Renderable 4D Try-on Proxy

2026-06-24 · Hao Sun, Hao Yan, Mengting Chen, Quanjian Song 외 arxiv

While Video Virtual Try-on (VVT) has achieved remarkable progress in synthesizing realistic garment overlays on dynamic subjects, existing paradigms remains fundamentally constrained by a passive dependency on source cam…

Virtual Try-on

Bridging 3D Gaussians and Semantic Occupancy for Comprehensive Open-Vocabulary Scene Understanding from Unposed Images

2026-07-02 · Hu Zhu, Bohan Li, Xianda Guo, Yanlun Peng 외 arxiv

Comprehensive 3D scene understanding from sparse, unposed images requires a model to recover renderable geometry, open-vocabulary semantics, and free/occupied 3D space without relying on external camera calibration. Rece…

Novel View SynthesisScene Understanding

Mono4DGS-HDR: High Dynamic Range 4D Gaussian Splatting from Alternating-exposure Monocular Videos

2025-10-21 · Jinfeng Liu, Lingtong Kong, Mi Zhou, Jinwen Chen 외 arxiv

We introduce Mono4DGS-HDR, the first system for reconstructing renderable 4D high dynamic range (HDR) scenes from unposed monocular low dynamic range (LDR) videos captured with alternating exposures. To tackle such a cha…

Video Reconstruction