paper-with-me

홈 › Papers

InstantSfM: Towards GPU-Native SfM for the Deep Learning Era

2025-10-15 · Jiankun Zhong, Zitong Zhan, Quankai Gao, Ziyu Chen, Haozhe Lou, Jiageng Mao, Ulrich Neumann, Chen Wang, Yue Wang arxiv

Structure-from-Motion (SfM) is a fundamental technique for recovering camera poses and scene structure from multi-view imagery, serving as a critical upstream component for applications ranging from 3D reconstruction to modern neural scene representations such as 3D Gaussian Splatting. However, most mature SfM systems remain CPU-centric and built upon traditional optimization toolchains, creating a growing mismatch with modern GPU-based, learning-driven pipelines and limiting scalability in large-scale scenes. While recent advances in GPU-accelerated bundle adjustment (BA) have demonstrated the potential of parallel sparse optimization, extending these techniques to build a complete global SfM system remains challenging due to unresolved issues in metric scale recovery and numerical robustness. In this paper, we implement a fully GPU-based and PyTorch-compatible global SfM system, named InstantSfM, to integrate seamlessly with modern learning pipelines. InstantSfM embeds metric depth priors directly into both global positioning and BA through a depth-constrained Jacobian structure, thereby resolving scale ambiguity within the optimization framework. To ensure numerical stability, we employ explicit filtering of under-constrained variables for the Jacobian matrix in an optimized GPU-friendly manner. Extensive experiments on diverse datasets demonstrate that InstantSfM achieves state-of-the-art efficiency while maintaining reconstruction accuracy comparable to both established classical pipelines and recent learning-based methods, showing up to ${\sim40\times}$ speedup over COLMAP on large-scale scenes.

📄 PDF Abstract BibTeX arXiv:2510.13310

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction

Similar Papers 제목 키워드 기반

On the Similarities Between Native, Non-native and Translated Texts

2016-09-11 · ACL 2016 8 · Ella Rabinovich, Sergiu Nisioi, Noam Ordan, Shuly Wintner

We present a computational analysis of three language varieties: native, advanced non-native, and translation. Our goal is to investigate the similarities and differences between non-native language productions and trans…

Translation

Detecting Spoof Voices in Asian Non-Native Speech: An Indonesian and Thai Case Study

2024-12-02 · Aulia Adila, Candy Olivia Mawalim, Masashi Unoki

This study focuses on building effective spoofing countermeasures (CMs) for non-native speech, specifically targeting Indonesian and Thai speakers. We constructed a dataset comprising both native and non-native speech to…

Non-native English lexicon creation for bilingual speech synthesis

2021-06-21 · Arun Baby, Pranav Jawale, Saranya Vinnaitherthan, Sumukh Badam 외

Bilingual English speakers speak English as one of their languages. Their English is of a non-native kind, and their conversations are of a code-mixed fashion. The intelligibility of a bilingual text-to-speech (TTS) syst…

Speech Synthesistext-to-speechText to Speech

Native Language Cognate Effects on Second Language Lexical Choice

2018-05-24 · TACL 2018 1 · Ella Rabinovich, Yulia Tsvetkov, Shuly Wintner

We present a computational analysis of cognate effects on the spontaneous linguistic productions of advanced non-native speakers. Introducing a large corpus of highly competent non-native English speakers, and using a se…

First-Passage Time Distributions in Two-State Protein Folding Kinetics: Exploring the Native-Like States vs Overcoming the Free Energy Barrier

2020-05-27 · Sergei F. Chekmarev

Using a beta-hairpin protein as a representative example of two-state folders, we studied how the exploration of native-like states affects the folding kinetics. It has been found that the first-passage time (FPT) distri…

Protein Folding