paper-with-me

Papers

GPVK-VL: Geometry-Preserving Virtual Keyframes for Visual Localization under Large Viewpoint Changes

2025-01-01 · CVPR 2025 1 · Yunxuan Li, Lei Fan, Xiaoying Xing, Jianxiong Zhou, Ying Wu

Visual localization, the task of determining the position and orientation of a camera, typically involves three core components: offline construction of a keyframe database, efficient online keyframes retrieval, and robust local feature matching. However, significant challenges arise when there are large viewpoint disparities between the query view and the database, such as attempting localization in a corridor previously build from an opposing direction. Intuitively, this issue can be addressed by synthesizing a set of virtual keyframes that cover all viewpoints. However, existing methods for synthesizing novel views to assist localization often fail to ensure geometric accuracy under large viewpoint changes. In this paper, we introduce a confidence-aware geometric prior into 2D Gaussian splatting to ensure the geometric accuracy of the scene. Then we can render novel views through the mesh with clear structures and accurate geometry, even under significant viewpoint changes, enabling the synthesis of a comprehensive set of virtual keyframes. Incorporating this geometry-preserving virtual keyframe database into the localization pipeline significantly enhances the robustness of visual localization.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Localization

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

LMap: Shape-Preserving Local Mappings for Biomedical Visualization

2018-09-17 · Saad Nadeem, Xianfeng GU, Arie Kaufman

Visualization of medical organs and biological structures is a challenging task because of their complex geometry and the resultant occlusions. Global spherical and planar mapping techniques simplify the complex geometry…

Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model

2026-04-10 · Shunkai Zhou, Zike Yan, Fei Xue, Dong Wu 외 arxiv

We present Online3R, a new sequential reconstruction framework that is capable of adapting to new scenes through online learning, effectively resolving inconsistency issues. Specifically, we introduce a set of learnable …

Self-Supervised Learning

I2V3D: Controllable image-to-video generation with 3D guidance

2025-03-12 · Zhiyuan Zhang, Dongdong Chen, Jing Liao

We present I2V3D, a novel framework for animating static images into dynamic videos with precise 3D control, leveraging the strengths of both 3D geometry guidance and advanced generative models. Our approach combines the…

3D geometryImage to Video GenerationVideo Generation

AutoScape: Geometry-Consistent Long-Horizon Scene Generation

2025-10-23 · Jiacheng Chen, Ziyu Jiang, Mingfu Liang, Bingbing Zhuang 외 arxiv

This paper proposes AutoScape, a long-horizon driving scene generation framework. At its core is a novel RGB-D diffusion model that iteratively generates sparse, geometrically consistent keyframes, serving as reliable an…

Scene GenerationPoint Clouds

Let Your Video Listen to Your Music!

2025-06-23 · Xinyu Zhang, Dong Gong, Zicheng Duan, Anton Van Den Hengel 외

Aligning the rhythm of visual motion in a video with a given music track is a practical need in multimedia production, yet remains an underexplored task in autonomous video editing. Effective alignment between motion and…

GPUMusic GenerationRhythmVideo Editing+1