paper-with-me

홈 › Papers

MoRE: 3D Visual Geometry Reconstruction Meets Mixture-of-Experts

2025-10-31 · Jingnan Gao, Zhe Wang, Xianze Fang, Xingyu Ren, Zhuo Chen, Shengqi Liu, Yuhao Cheng, Jiangjing Lyu, Xiaokang Yang, Yichao Yan arxiv

Recent advances in language and vision have demonstrated that scaling up model capacity consistently improves performance across diverse tasks. In 3D visual geometry reconstruction, large-scale training has likewise proven effective for learning versatile representations. However, further scaling of 3D models is challenging due to the complexity of geometric supervision and the diversity of 3D data. To overcome these limitations, we propose MoRE, a dense 3D visual foundation model based on a Mixture-of-Experts (MoE) architecture that dynamically routes features to task-specific experts, allowing them to specialize in complementary data aspects and enhance both scalability and adaptability. Aiming to improve robustness under real-world conditions, MoRE incorporates a confidence-based depth refinement module that stabilizes and refines geometric estimation. In addition, it integrates dense semantic features with globally aligned 3D backbone representations for high-fidelity surface normal prediction. MoRE is further optimized with tailored loss functions to ensure robust learning across diverse inputs and multiple geometric tasks. Extensive experiments demonstrate that MoRE achieves state-of-the-art performance across multiple benchmarks and supports effective downstream applications without extra computation.

📄 PDF Abstract BibTeX arXiv:2510.27234

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Geometry Meets Light: Leveraging Geometric Priors for Universal Photometric Stereo under Limited Multi-Illumination Cues

2025-11-17 · King-Man Tam, Satoshi Ikehata, Yuta Asano, Zhaoyi An 외 arxiv

Universal Photometric Stereo is a promising approach for recovering surface normals without strict lighting assumptions. However, it struggles when multi-illumination cues are unreliable, such as under biased lighting or…

3D Reconstruction

Pix2NPHM: Learning to Regress NPHM Reconstructions From a Single Image

2025-12-19 · Simon Giebenhain, Tobias Kirschstein, Liam Schoneveld, Davide Davoli 외 arxiv

Neural Parametric Head Models (NPHMs) are a recent advancement over mesh-based 3d morphable models (3DMMs) to facilitate high-fidelity geometric detail. However, fitting NPHMs to visual inputs is notoriously challenging …

SpaR3D-MoE: Adaptive 3D Spatial Reasoning from Sparse Views Meets Geometry-Inductive Mixture-of-Experts

2026-07-07 · Haida Feng, Hao Wei, Haolin Wang, Shiwei Li 외 arxiv

Recent Multimodal Large Language Models (MLLMs) struggle to bridge the representational gap between 2D semantic understanding and 3D spatial geometry. Existing 3D-aware models either rely on costly 3D-specific data or ut…

Spatial Reasoning

Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction

2026-01-07 · Jiaxin Huang, Yuanbo Yang, Bangbang Yang, Lin Ma 외 arxiv

We present Gen3R, a method that bridges the strong priors of foundational reconstruction models and video diffusion models for scene-level 3D generation. We repurpose the VGGT reconstruction model to produce geometric la…

Scene Generation3D GenerationPoint Clouds

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline

2025-08-06 · Linqing Zhao, Xiuwei Xu, Yirui Wang, Hao Wang 외 arxiv

Incrementally recovering real-sized 3D geometry from a pose-free RGB stream is a challenging task in 3D reconstruction, requiring minimal assumptions on input data. Existing methods can be broadly categorized into end-to…

3D ReconstructionPose Prediction