paper-with-me

홈 › Papers

FF3R: Feedforward Feature 3D Reconstruction from Unconstrained views

2026-04-10 · Chaoyi Zhou, Run Wang, Feng Luo, Mert D. Pesé, Zhiwen Fan, Yiqi Zhong, Siyu Huang arxiv

Recent advances in vision foundation models have revolutionized geometry reconstruction and semantic understanding. Yet, most of the existing approaches treat these capabilities in isolation, leading to redundant pipelines and compounded errors. This paper introduces FF3R, a fully annotation-free feed-forward framework that unifies geometric and semantic reasoning from unconstrained multi-view image sequences. Unlike previous methods, FF3R does not require camera poses, depth maps, or semantic labels, relying solely on rendering supervision for RGB and feature maps, establishing a scalable paradigm for unified 3D reasoning. In addition, we address two critical challenges in feedforward feature reconstruction pipelines, namely global semantic inconsistency and local structural inconsistency, through two key innovations: (i) a Token-wise Fusion Module that enriches geometry tokens with semantic context via cross-attention, and (ii) a Semantic-Geometry Mutual Boosting mechanism combining geometry-guided feature warping for global consistency with semantic-aware voxelization for local coherence. Extensive experiments on ScanNet and DL3DV-10K demonstrate FF3R's superior performance in novel-view synthesis, open-vocabulary semantic segmentation, and depth estimation, with strong generalization to in-the-wild scenarios, paving the way for embodied intelligence systems that demand both spatial and semantic understanding.

📄 PDF Abstract BibTeX arXiv:2604.09862

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Segmentation3D ReconstructionDepth Estimation

Similar Papers 제목 키워드 기반

WAT3R: Feedforward Underwater 3D Reconstruction

2026-07-23 · Jiayi Xu, Jiahao Lu, Ziqiang Zheng, Yihao Tan 외 arxiv

Reliable feedforward underwater 3D reconstruction remains challenging due to severe light attenuation and backscattering, which degrade visual quality and disrupt feature consistency across views, leading to inaccurate m…

Monocular Depth EstimationCamera Pose Estimation3D Reconstruction

RobustGS: Unified Boosting of Feedforward 3D Gaussian Splatting under Low-Quality Conditions

2025-08-05 · Anran Wu, Long Peng, Xin Di, Xueyuan Dai 외 arxiv

Feedforward 3D Gaussian Splatting (3DGS) overcomes the limitations of optimization-based 3DGS by enabling fast and high-quality reconstruction without the need for per-scene optimization. However, existing feedforward ap…

3D Reconstruction

UFO-4D: Unposed Feedforward 4D Reconstruction from Two Images

2026-02-27 · Junhwa Hur, Charles Herrmann, Songyou Peng, Philipp Henzler 외 arxiv

Dense 4D reconstruction from unposed images remains a critical challenge, with current methods relying on slow test-time optimization or fragmented, task-specific feedforward models. We introduce UFO-4D, a unified feedfo…

Camera Pose Estimation

ConFixGS: Learning to Fix Feedforward 3D Gaussian Splatting with Confidence-Aware Diffusion Priors in Driving Scenes

2026-05-10 · Rui Song, Tianhui Cai, Markus Gross, Xingcheng Zhou 외 arxiv

Feedforward 3D Gaussian Splatting (3DGS) often struggles in trajectory-based sparse-view driving scenes. Existing Gaussian repair methods mainly target optimization-based 3DGS, while diffusion-based repair is typically r…

Novel View Synthesis

Micro-macro Gaussian Splatting with Enhanced Scalability for Unconstrained Scene Reconstruction

2025-06-16 · Yihui Li, Chengxin Lv, Hongyu Yang, Di Huang

Reconstructing 3D scenes from unconstrained image collections poses significant challenges due to variations in appearance. In this paper, we propose Scalable Micro-macro Wavelet-based Gaussian Splatting (SMW-GS), a nove…

3D ReconstructionDiversity