paper-with-me

홈 › Papers

Large Point-to-Gaussian Model for Image-to-3D Generation

2024-08-20 · Longfei Lu, Huachen Gao, Tao Dai, Yaohua Zha, Zhi Hou, Junta Wu, Shu-Tao Xia

Recently, image-to-3D approaches have significantly advanced the generation quality and speed of 3D assets based on large reconstruction models, particularly 3D Gaussian reconstruction models. Existing large 3D Gaussian models directly map 2D image to 3D Gaussian parameters, while regressing 2D image to 3D Gaussian representations is challenging without 3D priors. In this paper, we propose a large Point-to-Gaussian model, that inputs the initial point cloud produced from large 3D diffusion model conditional on 2D image to generate the Gaussian parameters, for image-to-3D generation. The point cloud provides initial 3D geometry prior for Gaussian generation, thus significantly facilitating image-to-3D Generation. Moreover, we present the \textbf{A}ttention mechanism, \textbf{P}rojection mechanism, and \textbf{P}oint feature extractor, dubbed as \textbf{APP} block, for fusing the image features with point cloud features. The qualitative and quantitative experiments extensively demonstrate the effectiveness of the proposed approach on GSO and Objaverse datasets, and show the proposed method achieves state-of-the-art performance.

📄 PDF Abstract BibTeX arXiv:2408.10935

Code (0)

등록된 구현이 없습니다.

Tasks

3D Generation3D geometryImage to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

LucidDreamer: Domain-free Generation of 3D Gaussian Splatting Scenes

2023-11-22 · JaeYoung Chung, Suyoung Lee, Hyeongjin Nam, Jaerin Lee 외

With the widespread usage of VR devices and contents, demands for 3D scene generation techniques become more popular. Existing 3D scene generation models, however, limit the target scene to specific domain, primarily due…

Image GenerationScene Generation

GGS: Generalizable Gaussian Splatting for Lane Switching in Autonomous Driving

2024-09-04 · Huasong Han, Kaixuan Zhou, Xiaoxiao Long, Yusen Wang 외

We propose GGS, a Generalizable Gaussian Splatting method for Autonomous Driving which can achieve realistic rendering under large viewpoint changes. Previous generalizable 3D gaussian splatting methods are limited to re…

Autonomous DrivingDepth Estimation

PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image

2025-09-09 · Peng Li, Yisheng He, Yingdong Hu, Yuan Dong 외 arxiv

We present a feed-forward framework for Gaussian full-head synthesis from a single unposed image. Unlike previous work that relies on time-consuming GAN inversion and test-time optimization, our framework can reconstruct…

GaussianPainter: Painting Point Cloud into 3D Gaussians with Normal Guidance

2024-12-23 · Jingqiu Zhou, Lue Fan, Xuesong Chen, Linjiang Huang 외

In this paper, we present GaussianPainter, the first method to paint a point cloud into 3D Gaussians given a reference image. GaussianPainter introduces an innovative feed-forward approach to overcome the limitations of …

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization

2026-04-06 · Hongyu Liu, Xuan Wang, Zijian Wu, Yating Wang 외 arxiv

We introduce AvatarPointillist, a novel framework for generating dynamic 4D Gaussian avatars from a single portrait image. At the core of our method is a decoder-only Transformer that autoregressively generates a point c…