paper-with-me

홈 › Papers

Multi-Cali Anything: Dense Feature Multi-Frame Structure-from-Motion for Large-Scale Camera Array Calibration

2025-03-02 · Jinjiang You, Hewei Wang, Yijie Li, Mingxiao Huo, Long Van Tran Ha, Mingyuan Ma, Jinfeng Xu, Puzhen Wu, Shubham Garg, Wei Pu

Calibrating large-scale camera arrays, such as those in dome-based setups, is time-intensive and typically requires dedicated captures of known patterns. While extrinsics in such arrays are fixed due to the physical setup, intrinsics often vary across sessions due to factors like lens adjustments or temperature changes. In this paper, we propose a dense-feature-driven multi-frame calibration method that refines intrinsics directly from scene data, eliminating the necessity for additional calibration captures. Our approach enhances traditional Structure-from-Motion (SfM) pipelines by introducing an extrinsics regularization term to progressively align estimated extrinsics with ground-truth values, a dense feature reprojection term to reduce keypoint errors by minimizing reprojection loss in the feature space, and an intrinsics variance term for joint optimization across multiple frames. Experiments on the Multiface dataset show that our method achieves nearly the same precision as dedicated calibration processes, and significantly enhances intrinsics and 3D reconstruction accuracy. Fully compatible with existing SfM pipelines, our method provides an efficient and practical plug-and-play solution for large-scale camera setups. Our code is publicly available at: https://github.com/YJJfish/Multi-Cali-Anything

📄 PDF Abstract BibTeX arXiv:2503.00737

Code (1)

yjjfish/multi-cali-anything 공식 구현

Tasks

3D Reconstruction

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Semantic Localization Guiding Segment Anything Model For Reference Remote Sensing Image Segmentation

2025-06-12 · Shuyang Li, Shuang Wang, Zhuangzhuang Sun, Jing Xiao

The Reference Remote Sensing Image Segmentation (RRSIS) task generates segmentation masks for specified objects in images based on textual descriptions, which has attracted widespread attention and research interest. Cur…

Image SegmentationSegmentationSemantic SegmentationVisual Grounding

Count Anything

2026-05-29 · Mengqi Lei, Shuokun Cheng, Wei Bao, Shaoyi Du 외 arxiv

Object counting remains fragmented across domain-specific datasets and task formulations, despite rapid progress in generalist vision models. Existing counting models are often tailored to scenarios such as crowds, vehic…

Domain GeneralizationObject Counting

MapAnything: Universal Feed-Forward Metric 3D Reconstruction

2025-09-16 · Nikhil Keetha, Norman Müller, Johannes Schönberger, Lorenzo Porzi 외 arxiv

We introduce MapAnything, a unified transformer-based feed-forward model that ingests one or more images along with optional geometric inputs such as camera intrinsics, poses, depth, or partial reconstructions, and then …

Monocular Depth EstimationCamera Localization3D ReconstructionDepth Completion

MoonAnything: A Vision Benchmark with Large-Scale Lunar Supervised Data

2026-04-01 · Clémentine Grethen, Yuang Shi, Simone Gasparini, Géraldine Morin arxiv

Accurate perception of lunar surfaces is critical for modern lunar exploration missions. However, developing robust learning-based perception systems is hindered by the lack of datasets that provide both geometric and ph…

3D ReconstructionPose Estimation

AV-SAM: Segment Anything Model Meets Audio-Visual Localization and Segmentation

2023-05-03 · Shentong Mo, Yapeng Tian

Segment Anything Model (SAM) has recently shown its powerful effectiveness in visual segmentation tasks. However, there is less exploration concerning how SAM works on audio-visual tasks, such as visual sound localizatio…

DecoderObject LocalizationSegmentationVisual Localization