paper-with-me

홈 › Papers

BoostDream: Efficient Refining for High-Quality Text-to-3D Generation from Multi-View Diffusion

2024-01-30 · Yonghao Yu, Shunan Zhu, Huai Qin, Haorui Li

Witnessing the evolution of text-to-image diffusion models, significant strides have been made in text-to-3D generation. Currently, two primary paradigms dominate the field of text-to-3D: the feed-forward generation solutions, capable of swiftly producing 3D assets but often yielding coarse results, and the Score Distillation Sampling (SDS) based solutions, known for generating high-fidelity 3D assets albeit at a slower pace. The synergistic integration of these methods holds substantial promise for advancing 3D generation techniques. In this paper, we present BoostDream, a highly efficient plug-and-play 3D refining method designed to transform coarse 3D assets into high-quality. The BoostDream framework comprises three distinct processes: (1) We introduce 3D model distillation that fits differentiable representations from the 3D assets obtained through feed-forward generation. (2) A novel multi-view SDS loss is designed, which utilizes a multi-view aware 2D diffusion model to refine the 3D assets. (3) We propose to use prompt and multi-view consistent normal maps as guidance in refinement.Our extensive experiment is conducted on different differentiable 3D representations, revealing that BoostDream excels in generating high-quality 3D assets rapidly, overcoming the Janus problem compared to conventional SDS-based methods. This breakthrough signifies a substantial advancement in both the efficiency and quality of 3D generation processes.

📄 PDF Abstract BibTeX arXiv:2401.16764

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationText to 3D

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

Q-Refine: A Perceptual Quality Refiner for AI-Generated Image

2024-01-02 · Chunyi Li, HaoNing Wu, ZiCheng Zhang, Hongkun Hao 외

With the rapid evolution of the Text-to-Image (T2I) model in recent years, their unsatisfactory generation result has become a challenge. However, uniformly refining AI-Generated Images (AIGIs) of different qualities not…

Image Quality Assessment

AGFSync: Leveraging AI-Generated Feedback for Preference Optimization in Text-to-Image Generation

2024-03-20 · Jingkun An, Yinghao Zhu, Zongjian Li, Enshen Zhou 외

Text-to-Image (T2I) diffusion models have achieved remarkable success in image generation. Despite their progress, challenges remain in both prompt-following ability, image quality and lack of high-quality datasets, whic…

Image GenerationText to Image GenerationText-to-Image GenerationVisual Question Answering (VQA)

Text2Immersion: Generative Immersive Scene with 3D Gaussians

2023-12-14 · Hao Ouyang, Kathryn Heal, Stephen Lombardi, Tiancheng Sun

We introduce Text2Immersion, an elegant method for producing high-quality 3D immersive scenes from text prompts. Our proposed pipeline initiates by progressively generating a Gaussian cloud using pre-trained 2D diffusion…

Depth EstimationDiversityScene Generation

GradeADreamer: Enhanced Text-to-3D Generation Using Gaussian Splatting and Multi-View Diffusion

2024-06-14 · Trapoom Ukarapol, Kevin Pruvost

Text-to-3D generation has shown promising results, yet common challenges such as the Multi-face Janus problem and extended generation time for high-quality assets. In this paper, we address these issues by introducing a …

3D GenerationGPUText to 3D

Refining Visual Artifacts in Diffusion Models via Explainable AI-based Flaw Activation Maps

2025-12-09 · Seoyeon Lee, Gwangyeol Yu, Chaewon Kim, Jonghyuk Park arxiv

Diffusion models have achieved remarkable success in image synthesis. However, addressing artifacts and unrealistic regions remains a critical challenge. We propose self-refining diffusion, a novel framework that enhance…

Text-to-Image Generation