paper-with-me

홈 › Papers

Magic3D: High-Resolution Text-to-3D Content Creation

2022-11-18 · CVPR 2023 1 · Chen-Hsuan Lin, Jun Gao, Luming Tang, Towaki Takikawa, Xiaohui Zeng, Xun Huang, Karsten Kreis, Sanja Fidler, Ming-Yu Liu, Tsung-Yi Lin

DreamFusion has recently demonstrated the utility of a pre-trained text-to-image diffusion model to optimize Neural Radiance Fields (NeRF), achieving remarkable text-to-3D synthesis results. However, the method has two inherent limitations: (a) extremely slow optimization of NeRF and (b) low-resolution image space supervision on NeRF, leading to low-quality 3D models with a long processing time. In this paper, we address these limitations by utilizing a two-stage optimization framework. First, we obtain a coarse model using a low-resolution diffusion prior and accelerate with a sparse 3D hash grid structure. Using the coarse representation as the initialization, we further optimize a textured 3D mesh model with an efficient differentiable renderer interacting with a high-resolution latent diffusion model. Our method, dubbed Magic3D, can create high quality 3D mesh models in 40 minutes, which is 2x faster than DreamFusion (reportedly taking 1.5 hours on average), while also achieving higher resolution. User studies show 61.7% raters to prefer our approach over DreamFusion. Together with the image-conditioned generation capabilities, we provide users with new ways to control 3D synthesis, opening up new avenues to various creative applications.

📄 PDF Abstract BibTeX arXiv:2211.10440

Code (1)

chinhsuanwu/dreamfusionacc pytorch

Tasks

NeRFText to 3DVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

3D-CLFusion: Fast Text-to-3D Rendering with Contrastive Latent Diffusion

2023-03-21 · Yu-Jhe Li, Tao Xu, Ji Hou, Bichen Wu 외

We tackle the task of text-to-3D creation with pre-trained latent-based NeRFs (NeRFs that generate 3D objects given input latent code). Recent works such as DreamFusion and Magic3D have shown great success in generating …

Contrastive LearningNeRFText to 3D

Z-Magic: Zero-shot Multiple Attributes Guided Image Creator

2025-01-01 · CVPR 2025 1 · Yingying Deng, Xiangyu He, Fan Tang, WeiMing Dong

The customization of multiple attributes has gained increasing popularity with the rising demand for personalized content creation. Despite promising empirical results, the contextual coherence between different attr…

AttributeImage GenerationMulti-Task Learning

Magic-Me: Identity-Specific Video Customized Diffusion

2024-02-14 · Ze Ma, Daquan Zhou, Chun-Hsiao Yeh, Xue-She Wang 외

Creating content with specified identities (ID) has attracted significant interest in the field of generative models. In the field of text-to-image generation (T2I), subject-driven creation has achieved great progress wi…

Image GenerationText to Image GenerationText-to-Image GenerationVideo Generation

Magic123: One Image to High-Quality 3D Object Generation Using Both 2D and 3D Diffusion Priors

2023-06-30 · Guocheng Qian, Jinjie Mai, Abdullah Hamdi, Jian Ren 외

We present Magic123, a two-stage coarse-to-fine approach for high-quality, textured 3D meshes generation from a single unposed image in the wild using both2D and 3D priors. In the first stage, we optimize a neural radian…

Image to 3D

MagicFight: Personalized Martial Arts Combat Video Generation

2026-01-05 · Jiancheng Huang, Mingfu Yan, Songyan Chen, Yi Huang 외 arxiv

Amid the surge in generic text-to-video generation, the field of personalized human video generation has witnessed notable advancements, primarily concentrated on single-person scenarios. However, to our knowledge, the d…

Text-to-Video Generation