paper-with-me

Papers

LucidFusion: Generating 3D Gaussians with Arbitrary Unposed Images

2024-10-21 · Hao He, Yixun Liang, Luozhou Wang, Yuanhao Cai, Xinli Xu, Hao-Xiang Guo, Xiang Wen, Yingcong Chen

Recent large reconstruction models have made notable progress in generating high-quality 3D objects from single images. However, these methods often struggle with controllability, as they lack information from multiple views, leading to incomplete or inconsistent 3D reconstructions. To address this limitation, we introduce LucidFusion, a flexible end-to-end feed-forward framework that leverages the Relative Coordinate Map (RCM). Unlike traditional methods linking images to 3D world thorough pose, LucidFusion utilizes RCM to align geometric features coherently across different views, making it highly adaptable for 3D generation from arbitrary, unposed images. Furthermore, LucidFusion seamlessly integrates with the original single-image-to-3D pipeline, producing detailed 3D Gaussians at a resolution of $512 \times 512$, making it well-suited for a wide range of applications.

📄 PDF Abstract BibTeX arXiv:2410.15636

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationImage to 3D

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth Priors

2026-03-24 · Chuanqing Zhuang, Xin Lu, Zehui Deng, Zhengda Lu 외 arxiv

Omnidirectional 3D Gaussian Splatting with panoramas is a key technique for 3D scene representation, and existing methods typically rely on slow SfM to provide camera poses and sparse points priors. In this work, we prop…

Camera Pose EstimationNovel View Synthesis

Unposed 3DGS Reconstruction with Probabilistic Procrustes Mapping

2025-07-24 · Chong Cheng, Zijian Wang, Sicheng Yu, Yu Hu 외 arxiv

3D Gaussian Splatting (3DGS) has emerged as a core technique for 3D representation. Its effectiveness largely depends on precise camera poses and accurate point cloud initialization, which are often derived from pretrain…

Pose EstimationPoint Clouds

RegGS: Unposed Sparse Views Gaussian Splatting with 3DGS Registration

2025-07-10 · Chong Cheng, Yu Hu, Sicheng Yu, Beizhen Zhao 외 arxiv

3D Gaussian Splatting (3DGS) has demonstrated its potential in reconstructing scenes from unposed images. However, optimization-based 3DGS methods struggle with sparse views due to limited prior knowledge. Meanwhile, fee…

Pose Estimation

GaussiGAN: Controllable Image Synthesis with 3D Gaussians from Unposed Silhouettes

2021-06-24 · Youssef A. Mejjati, Isa Milefchik, Aaron Gokaslan, Oliver Wang 외

We present an algorithm that learns a coarse 3D representation of objects from unposed multi-view 2D mask supervision, then uses it to generate detailed mask and image texture. In contrast to existing voxel-based methods…

Image GenerationObjectObject Reconstruction

Bridging 3D Gaussians and Semantic Occupancy for Comprehensive Open-Vocabulary Scene Understanding from Unposed Images

2026-07-02 · Hu Zhu, Bohan Li, Xianda Guo, Yanlun Peng 외 arxiv

Comprehensive 3D scene understanding from sparse, unposed images requires a model to recover renderable geometry, open-vocabulary semantics, and free/occupied 3D space without relying on external camera calibration. Rece…

Novel View SynthesisScene Understanding