paper-with-me

홈 › Papers

Cross-Modal 3D Shape Generation and Manipulation

2022-07-24 · Zezhou Cheng, Menglei Chai, Jian Ren, Hsin-Ying Lee, Kyle Olszewski, Zeng Huang, Subhransu Maji, Sergey Tulyakov

Creating and editing the shape and color of 3D objects require tremendous human effort and expertise. Compared to direct manipulation in 3D interfaces, 2D interactions such as sketches and scribbles are usually much more natural and intuitive for the users. In this paper, we propose a generic multi-modal generative model that couples the 2D modalities and implicit 3D representations through shared latent spaces. With the proposed model, versatile 3D generation and manipulation are enabled by simply propagating the editing from a specific 2D controlling modality through the latent spaces. For example, editing the 3D shape by drawing a sketch, re-colorizing the 3D surface via painting color scribbles on the 2D rendering, or generating 3D shapes of a certain category given one or a few reference images. Unlike prior works, our model does not require re-training or fine-tuning per editing task and is also conceptually simple, easy to implement, robust to input domain shifts, and flexible to diverse reconstruction on partial 2D inputs. We evaluate our framework on two representative 2D modalities of grayscale line sketches and rendered color images, and demonstrate that our method enables various shape manipulation and generation tasks with these 2D modalities.

📄 PDF Abstract BibTeX arXiv:2207.11795

Code (0)

등록된 구현이 없습니다.

Tasks

3D Generation3D Shape Generation

Similar Papers 제목 키워드 기반

ShapeScaffolder: Structure-Aware 3D Shape Generation from Text

2023-01-01 · ICCV 2023 1 · Xi Tian, Yong-Liang Yang, Qi Wu

We present ShapeScaffolder, a structure-based neural network for generating colored 3D shapes based on text input. The approach, similar to providing scaffolds as internal structural supports and adding more details …

3D Shape GenerationText Matching

Text2Face: A Multi-Modal 3D Face Model

2023-03-05 · Will Rowan, Patrik Huber, Nick Pears, Andrew Keeling

We present the first 3D morphable modelling approach, whereby 3D face shape can be directly and completely defined using a textual prompt. Building on work in multi-modal learning, we extend the FLAME head model to a com…

Face Modelmodel

DRAW2ACT: Turning Depth-Encoded Trajectories into Robotic Demonstration Videos

2025-12-16 · Yang Bai, Liudi Yang, George Eskandar, Fengyi Shen 외 arxiv

Video diffusion models provide powerful real-world simulators for embodied AI but remain limited in controllability for robotic manipulation. Recent works on trajectory-conditioned video generation address this gap but o…

Video Generation

Neural Wavelet-domain Diffusion for 3D Shape Generation, Inversion, and Manipulation

2023-02-01 · Jingyu Hu, Ka-Hei Hui, Zhengzhe Liu, Ruihui Li 외

This paper presents a new approach for 3D shape generation, inversion, and manipulation, through a direct generative modeling on a continuous implicit representation in wavelet domain. Specifically, we propose a compact …

3D Shape Generation

Multimodal perception for dexterous manipulation

2021-12-28 · Guanqun Cao, Shan Luo

Humans usually perceive the world in a multimodal way that vision, touch, sound are utilised to understand surroundings from various dimensions. These senses are combined together to achieve a synergistic effect where th…

3D ReconstructionFrictionTranslation