paper-with-me

Papers

Progressive3D: Progressively Local Editing for Text-to-3D Content Creation with Complex Semantic Prompts

2023-10-18 · Xinhua Cheng, Tianyu Yang, Jianan Wang, Yu Li, Lei Zhang, Jian Zhang, Li Yuan

Recent text-to-3D generation methods achieve impressive 3D content creation capacity thanks to the advances in image diffusion models and optimizing strategies. However, current methods struggle to generate correct 3D content for a complex prompt in semantics, i.e., a prompt describing multiple interacted objects binding with different attributes. In this work, we propose a general framework named Progressive3D, which decomposes the entire generation into a series of locally progressive editing steps to create precise 3D content for complex prompts, and we constrain the content change to only occur in regions determined by user-defined region prompts in each editing step. Furthermore, we propose an overlapped semantic component suppression technique to encourage the optimization process to focus more on the semantic differences between prompts. Extensive experiments demonstrate that the proposed Progressive3D framework generates precise 3D content for prompts with complex semantics and is general for various text-to-3D methods driven by different 3D representations.

📄 PDF Abstract BibTeX arXiv:2310.11784

Code (0)

등록된 구현이 없습니다.

Tasks

3D GenerationText to 3D

Methods 이 논문이 사용한 방법론

Focus 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

ReVideo: Remake a Video with Motion and Content Control

2024-05-22 · Chong Mou, Mingdeng Cao, Xintao Wang, Zhaoyang Zhang 외

Despite significant advancements in video generation and editing using diffusion models, achieving accurate and localized video editing remains a substantial challenge. Additionally, most existing video editing methods p…

Video EditingVideo Generation

3DGS-Drag: Dragging Gaussians for Intuitive Point-Based 3D Editing

2026-01-12 · Jiahua Dong, Yu-Xiong Wang arxiv

The transformative potential of 3D content creation has been progressively unlocked through advancements in generative models. Recently, intuitive drag editing with geometric changes has attracted significant attention i…

Coin3D: Controllable and Interactive 3D Assets Generation with Proxy-Guided Conditioning

2024-05-13 · Wenqi Dong, Bangbang Yang, Lin Ma, Xiao Liu 외

As humans, we aspire to create media content that is both freely willed and readily controlled. Thanks to the prominent development of generative techniques, we now can easily utilize 2D diffusion methods to synthesize i…

3D Generation

MEDIC: Zero-shot Music Editing with Disentangled Inversion Control

2024-07-18 · Huadai Liu, Jialei Wang, Xiangtai Li, Rongjie Huang 외

Text-guided diffusion models make a paradigm shift in audio generation, facilitating the adaptability of source audio to conform to specific textual prompts. Recent works introduce inversion techniques, like DDIM inversi…

Audio Generation

ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks

2026-04-09 · Jiayang Xu, Fan Zhuo, Majun Zhang, Changhao Pan 외 arxiv

Current video editing models often rely on expensive paired video data, which limits their practical scalability. In essence, most video editing tasks can be formulated as a decoupled spatiotemporal process, where the te…