paper-with-me

홈 › Papers

ShapeScaffolder: Structure-Aware 3D Shape Generation from Text

2023-01-01 · ICCV 2023 1 · Xi Tian, Yong-Liang Yang, Qi Wu

We present ShapeScaffolder, a structure-based neural network for generating colored 3D shapes based on text input. The approach, similar to providing scaffolds as internal structural supports and adding more details to them, aims to capture finer text-shape connections and improve the quality of generated shapes. Traditional text-to-shape methods often generate 3D shapes as a whole. However, humans tend to understand both shape and text as being structure-based. For example, a table is interpreted as being composed of legs, a seat, and a back; similarly, texts possess inherent linguistic structures that can be analyzed as dependency graphs, depicting the relationships between entities within the text. We believe structure-aware shape generation can bring finer text-shape connections and improve shape generation quality. However, the lack of explicit shape structure and the high freedom of text structure make cross-modality learning challenging. To address these challenges, we first build the structured shape implicit fields in an unsupervised manner. We then propose the part-level attention mechanism between shape parts and textual graph nodes to align the two modalities at the structural level. Finally, we employ a shape refiner to add further detail to the predicted structure, yielding the final results. Extensive experimentation demonstrates that our approaches outperform state-of-the-art methods in terms of both shape fidelity and shape-text matching. Our methods also allow for part-level manipulation and improved part-level completeness.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3D Shape GenerationText Matching

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Part-aware Shape Generation with Latent 3D Diffusion of Neural Voxel Fields

2024-05-02 · Yuhang Huang, SHilong Zou, Xinwang Liu, Kai Xu

This paper presents a novel latent 3D diffusion model for the generation of neural voxel fields, aiming to achieve accurate part-aware structures. Compared to existing methods, there are two key designs to ensure high-qu…

Decoder

NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and Generation

2024-03-27 · Ruikai Cui, Weizhe Liu, Weixuan Sun, Senbo Wang 외

3D shape generation aims to produce innovative 3D content adhering to specific conditions and constraints. Existing methods often decompose 3D shapes into a sequence of localized components, treating each element in isol…

3D Shape Generation3D Shape Modeling

SP-GAN: Sphere-Guided 3D Shape Generation and Manipulation

2021-08-10 · Ruihui Li, Xianzhi Li, Ka-Hei Hui, Chi-Wing Fu

We present SP-GAN, a new unsupervised sphere-guided generative model for direct synthesis of 3D shapes in the form of point clouds. Compared with existing models, SP-GAN is able to synthesize diverse and high-quality sha…

3D Shape Generation

StrucADT: Generating Structure-controlled 3D Point Clouds with Adjacency Diffusion Transformer

2025-09-28 · Zhenyu Shu, Jiajun Shen, Zhongui Chen, Xiaoguang Han 외 arxiv

In the field of 3D point cloud generation, numerous 3D generative models have demonstrated the ability to generate diverse and realistic 3D shapes. However, the majority of these approaches struggle to generate controlla…

Point Cloud GenerationPoint Clouds

ShapeWords: Guiding Text-to-Image Synthesis with 3D Shape-Aware Prompts

2024-12-03 · CVPR 2025 1 · Dmitry Petrov, Pradyumn Goyal, Divyansh Shivashok, Yuanming Tao 외

We introduce ShapeWords, an approach for synthesizing images based on 3D shape guidance and text prompts. ShapeWords incorporates target 3D shape information within specialized tokens embedded together with the input tex…

Image Generation