paper-with-me

홈 › Papers

TextMesh: Generation of Realistic 3D Meshes From Text Prompts

2023-04-24 · Christina Tsalicoglou, Fabian Manhardt, Alessio Tonioni, Michael Niemeyer, Federico Tombari

The ability to generate highly realistic 2D images from mere text prompts has recently made huge progress in terms of speed and quality, thanks to the advent of image diffusion models. Naturally, the question arises if this can be also achieved in the generation of 3D content from such text prompts. To this end, a new line of methods recently emerged trying to harness diffusion models, trained on 2D images, for supervision of 3D model generation using view dependent prompts. While achieving impressive results, these methods, however, have two major drawbacks. First, rather than commonly used 3D meshes, they instead generate neural radiance fields (NeRFs), making them impractical for most real applications. Second, these approaches tend to produce over-saturated models, giving the output a cartoonish looking effect. Therefore, in this work we propose a novel method for generation of highly realistic-looking 3D meshes. To this end, we extend NeRF to employ an SDF backbone, leading to improved 3D mesh extraction. In addition, we propose a novel way to finetune the mesh texture, removing the effect of high saturation and improving the details of the output 3D mesh.

📄 PDF Abstract BibTeX arXiv:2304.12439

Code (1)

threestudio-project/threestudio jax

Tasks

NeRF

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Text-guided High-definition Consistency Texture Model

2023-05-10 · Zhibin Tang, Tiantong He

With the advent of depth-to-image diffusion models, text-guided generation, editing, and transfer of realistic textures are no longer difficult. However, due to the limitations of pre-trained diffusion models, they can o…

modelparameter-efficient fine-tuningtext-guided-generationVocal Bursts Intensity Prediction

LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models

2024-11-14 · Zhengyi Wang, Jonathan Lorraine, Yikai Wang, Hang Su 외

This work explores expanding the capabilities of large language models (LLMs) pretrained on text to generate 3D meshes within a unified model. This offers key advantages of (1) leveraging spatial knowledge already embedd…

3D GenerationText Generation

Open-Universe Indoor Scene Generation using LLM Program Synthesis and Uncurated Object Databases

2024-02-05 · Rio Aguina-Kang, Maxim Gumin, Do Heon Han, Stewart Morris 외

We present a system for generating indoor scenes in response to text prompts. The prompts are not limited to a fixed vocabulary of scene descriptions, and the objects in generated scenes are not restricted to a fixed set…

Layout GenerationObjectProgram SynthesisScene Generation+1

THOM: Generating Physically Plausible Hand-Object Meshes From Text

2026-04-03 · Uyoung Jeong, Yihalem Yimolal Tiruneh, Hyung Jin Chang, Seungryul Baek 외 arxiv

Generating photorealistic 3D hand-object interactions (HOIs) from text is important for applications like robotic grasping and AR/VR content creation. In practice, however, achieving both visual fidelity and physical pla…

Robotic Grasping

ClipMatrix: Text-controlled Creation of 3D Textured Meshes

2021-09-27 · Nikolay Jetchev

If a picture is worth thousand words, a moving 3d shape must be worth a million. We build upon the success of recent generative methods that create images fitting the semantics of a text prompt, and extend it to the cont…