paper-with-me

Camera Pose Estimation

1개 벤치마크 · 논문 408편 · 이 태스크의 논문 보기 →

Benchmarks

Most implemented

CubeSLAM: Monocular 3D Object SLAM

2018-06-01 · 구현 5개

Event-based Stereo Visual Odometry

2020-07-30 · 구현 4개

Papers

GS-CPE: Unified 6-Degree-of-Freedom Camera Pose Estimation via 3D Gaussian Splatting

2026-08-11 · Huaiyuan Weng, Chul Min Yeum, Su-Min Kang arxiv

Despite substantial progress in visual localization, from scene coordinate regression to direct camera pose regression, achieving both robust generalization and high accuracy remain challenging. This study introduces GS-…

Camera Pose EstimationVisual Localization

WAT3R: Feedforward Underwater 3D Reconstruction

2026-07-23 · Jiayi Xu, Jiahao Lu, Ziqiang Zheng, Yihao Tan 외 arxiv

Reliable feedforward underwater 3D reconstruction remains challenging due to severe light attenuation and backscattering, which degrade visual quality and disrupt feature consistency across views, leading to inaccurate m…

Monocular Depth EstimationCamera Pose Estimation3D Reconstruction

OmniX: Any-view and Any-time 4D Reconstruction via Feed-forward Trajectory Fields

2026-07-12 · Yanqin Jiang, Tengfei Wang, Zhengwei Wang, Chenjie Cao 외 arxiv

Previous feed-forward 4D reconstruction methods either predict per-frame static point clouds, ignoring foreground motion, or estimate point cloud trajectories while being limited to small camera motions. This restricts t…

Camera Pose EstimationTrajectory PredictionDepth EstimationPoint Tracking

Video Generation Models are General-Purpose Vision Learners

2026-07-10 · Letian Wang, Chuhan Zhang, Rishabh Kabra, Jasper Uijlings 외 arxiv

Driven by next-token prediction, NLP shifted from task-specific models into powerful generalist foundation models. What, then, is the equivalent catalyst needed to achieve a general-purpose model in computer vision? In t…

Text-to-Video GenerationCamera Pose Estimation

NoDrift3R: Raymap-Guided Coupling for Drift-Robust Unposed Feed-Forward 3D Reconstruction

2026-07-08 · Xiangyu Sun, Liu Liu, Seungkwon Yang, Jingbing Han 외 arxiv

Pose-Free Feed-forward 3D Gaussian Splatting (3DGS) has recently emerged as a powerful paradigm for fast scene reconstruction. However, its performance degrades significantly in long image sequences due to cumulative cam…

Camera Pose Estimation3D Reconstruction

Vision as Unified Multimodal Generation

2026-07-07 · Xiaoyang Han, Jianhua Li, Kewang Deng, Zukai Chen 외 arxiv

We formulate computer vision as unified multimodal generation, where heterogeneous visual tasks are expressed in the native text and image generation spaces of a unified multimodal model, without task-specific architectu…

Camera Pose Estimationmultimodal generationDepth EstimationImage Generation

전체 408편 보기 →