paper-with-me

Papers

AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer

2026-02-12 · Lingting Zhu, Shengju Qian, Haidi Fan, Jiayu Dong, Zhenchao Jin, Siwei Zhou, Gen Dong, Xin Wang, Lequan Yu arxiv

The digital industry demands high-quality, diverse modular 3D assets, especially for user-generated content~(UGC). In this work, we introduce AssetFormer, an autoregressive Transformer-based model designed to generate modular 3D assets from textual descriptions. Our pilot study leverages real-world modular assets collected from online platforms. AssetFormer tackles the challenge of creating assets composed of primitives that adhere to constrained design parameters for various applications. By innovatively adapting module sequencing and decoding techniques inspired by language models, our approach enhances asset generation quality through autoregressive modeling. Initial results indicate the effectiveness of AssetFormer in streamlining asset creation for professional development and UGC scenarios. This work presents a flexible framework extendable to various types of modular 3D assets, contributing to the broader field of 3D content generation. The code is available at https://github.com/Advocate99/AssetFormer.

📄 PDF Abstract BibTeX arXiv:2602.12100

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers

2024-06-14 · YiWen Chen, Tong He, Di Huang, Weicai Ye 외

Recently, 3D assets created via reconstruction and generation have matched the quality of manually crafted assets, highlighting their potential for replacement. However, this potential is largely unrealized because these…

Decoder

MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents

2026-06-01 · Minkyung Kwon, Jinhyeok Choi, Youngjin Shin, Jaeyeong Kim 외 arxiv

We present MORPHOS, a novel autoregressive framework that generates dynamic 3D assets from videos across diverse representations, including meshes, 3D Gaussians, and radiance fields. Existing methods are typically limite…

TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction

2024-12-22 · Xuying Zhang, Yutong Liu, Yangguang Li, Renrui Zhang 외

We present TAR3D, a novel framework that consists of a 3D-aware Vector Quantized-Variational AutoEncoder (VQ-VAE) and a Generative Pre-trained Transformer (GPT) to generate high-quality 3D assets. The core insight of thi…

Image to 3DText to 3D

SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control

2022-10-31 · Xiaochuang Han, Sachin Kumar, Yulia Tsvetkov

Despite the growing success of diffusion models in continuous-valued domains (e.g., images), similar efforts for discrete domains such as text have yet to match the performance of autoregressive language models. In this …

DiversityLanguage ModelingLanguage ModellingText Generation

MeshArt: Generating Articulated Meshes with Structure-Guided Transformers

2024-12-16 · CVPR 2025 1 · Daoyi Gao, Yawar Siddiqui, Lei LI, Angela Dai

Articulated 3D object generation is fundamental for creating realistic, functional, and interactable virtual assets which are not simply static. We introduce MeshArt, a hierarchical transformer-based approach to generate…

Object