paper-with-me

Papers

Control3Diff: Learning Controllable 3D Diffusion Models from Single-view Images

2023-04-13 · Jiatao Gu, Qingzhe Gao, Shuangfei Zhai, Baoquan Chen, Lingjie Liu, Josh Susskind

Diffusion models have recently become the de-facto approach for generative modeling in the 2D domain. However, extending diffusion models to 3D is challenging due to the difficulties in acquiring 3D ground truth data for training. On the other hand, 3D GANs that integrate implicit 3D representations into GANs have shown remarkable 3D-aware generation when trained only on single-view image datasets. However, 3D GANs do not provide straightforward ways to precisely control image synthesis. To address these challenges, We present Control3Diff, a 3D diffusion model that combines the strengths of diffusion models and 3D GANs for versatile, controllable 3D-aware image synthesis for single-view datasets. Control3Diff explicitly models the underlying latent distribution (optionally conditioned on external inputs), thus enabling direct control during the diffusion process. Moreover, our approach is general and applicable to any type of controlling input, allowing us to train it with the same diffusion objective without any auxiliary supervision. We validate the efficacy of Control3Diff on standard image generation benchmarks, including FFHQ, AFHQ, and ShapeNet, using various conditioning inputs such as images, sketches, and text prompts. Please see the project website (\url{https://jiataogu.me/control3diff}) for video comparisons.

📄 PDF Abstract BibTeX arXiv:2304.06700

Code (0)

등록된 구현이 없습니다.

Tasks

3D-Aware Image SynthesisImage Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DreamComposer++: Empowering Diffusion Models with Multi-View Conditions for 3D Content Generation

2025-07-03 · Yunhan Yang, Shuo Chen, Yukun Huang, Xiaoyang Wu 외 arxiv

Recent advancements in leveraging pre-trained 2D diffusion models achieve the generation of high-quality novel views from a single in-the-wild image. However, existing works face challenges in producing controllable nove…

3D Object ReconstructionNovel View Synthesis

DreamComposer: Controllable 3D Object Generation via Multi-View Conditions

2023-12-06 · CVPR 2024 1 · Yunhan Yang, Yukun Huang, Xiaoyang Wu, Yuan-Chen Guo 외

Utilizing pre-trained 2D large-scale generative models, recent works are capable of generating high-quality novel views from a single in-the-wild image. However, due to the lack of information from multiple views, these …

3D Object ReconstructionNovel View SynthesisObjectObject Reconstruction

MVControl: Adding Conditional Control to Multi-view Diffusion for Controllable Text-to-3D Generation

2023-11-24 · Zhiqi Li, Yiming Chen, Lingzhe Zhao, Peidong Liu

We introduce MVControl, a novel neural network architecture that enhances existing pre-trained multi-view 2D diffusion models by incorporating additional input conditions, e.g. edge maps. Our approach enables the generat…

3D GenerationImage GenerationText to 3D

Points-to-3D: Bridging the Gap between Sparse Points and Shape-Controllable Text-to-3D Generation

2023-07-26 · Chaohui Yu, Qiang Zhou, Jingliang Li, Zhe Zhang 외

Text-to-3D generation has recently garnered significant attention, fueled by 2D diffusion models trained on billions of image-text pairs. Existing methods primarily rely on score distillation to leverage the 2D diffusion…

3D GenerationNeRFText to 3D

Controllable Generation with Text-to-Image Diffusion Models: A Survey

2024-03-07 · Pu Cao, Feng Zhou, Qing Song, Lu Yang

In the rapidly advancing realm of visual generation, diffusion models have revolutionized the landscape, marking a significant shift in capabilities with their impressive text-guided generative functions. However, relyin…

Denoising