paper-with-me

Papers

AToM: Amortized Text-to-Mesh using 2D Diffusion

2024-02-01 · Guocheng Qian, Junli Cao, Aliaksandr Siarohin, Yash Kant, Chaoyang Wang, Michael Vasilkovsky, Hsin-Ying Lee, Yuwei Fang, Ivan Skorokhodov, Peiye Zhuang, Igor Gilitschenski, Jian Ren, Bernard Ghanem, Kfir Aberman, Sergey Tulyakov

We introduce Amortized Text-to-Mesh (AToM), a feed-forward text-to-mesh framework optimized across multiple text prompts simultaneously. In contrast to existing text-to-3D methods that often entail time-consuming per-prompt optimization and commonly output representations other than polygonal meshes, AToM directly generates high-quality textured meshes in less than 1 second with around 10 times reduction in the training cost, and generalizes to unseen prompts. Our key idea is a novel triplane-based text-to-mesh architecture with a two-stage amortized optimization strategy that ensures stable training and enables scalability. Through extensive experiments on various prompt benchmarks, AToM significantly outperforms state-of-the-art amortized approaches with over 4 times higher accuracy (in DF415 dataset) and produces more distinguishable and higher-quality 3D outputs. AToM demonstrates strong generalizability, offering finegrained 3D assets for unseen interpolated prompts without further optimization during inference, unlike per-prompt solutions.

📄 PDF Abstract BibTeX arXiv:2402.00867

Code (0)

등록된 구현이 없습니다.

Tasks

Text to 3D

Similar Papers 제목 키워드 기반

3D Cardiac Anatomy Generation Using Mesh Latent Diffusion Models

2025-08-18 · Jolanta Mozyrska, Marcel Beetz, Luke Melas-Kyriazi, Vicente Grau 외 arxiv

Diffusion models have recently gained immense interest for their generative capabilities, specifically the high quality and diversity of the synthesized data. However, examples of their applications in 3D medical imaging…

Conditional Latent Diffusion Model with Fourier-based Motion Modelling for Virtual Population Synthesis

2026-06-02 · Shaokun Lan, Haoran Dou, Jinghan Huang, Arezoo Zakeri 외 arxiv

In-silico trials of medical devices require the generation of virtual populations of anatomies. In cardiovascular applications, virtual anatomy is typically represented as a 3D+t mesh sampled from a generative model. How…

LATTE3D: Large-scale Amortized Text-To-Enhanced3D Synthesis

2024-03-22 · Kevin Xie, Jonathan Lorraine, Tianshi Cao, Jun Gao 외

Recent text-to-3D generation approaches produce impressive 3D results but require time-consuming optimization that can take up to an hour per prompt. Amortized methods like ATT3D optimize multiple prompts simultaneously …

3D GenerationText to 3D

High-Fidelity Medical Shape Generation via Skeletal Latent Diffusion

2026-03-08 · Guoqing Zhang, Jingyun Yang, Siqi Chen, Anping Zhang 외 arxiv

Anatomy shape modeling is a fundamental problem in medical data analysis. However, the geometric complexity and topological variability of anatomical structures pose significant challenges to accurate anatomical shape ge…

Computational EfficiencyPoint Clouds

DIPHINE: Diffusion-based $Φ$-ID Neural Estimator

2026-06-17 · Simon Pedro Galeano Munoz, Mustapha Bounoua, Giulio Franzese, Pietro Michiardi 외 arxiv

Uncovering the true informational architecture of real-world complex systems requires disentangling how their components uniquely store, redundantly share, and synergistically integrate information over time. Integrated …