paper-with-me

홈 › Papers

InfiniHuman: Infinite 3D Human Creation with Precise Control

2025-10-13 · Yuxuan Xue, Xianghui Xie, Margaret Kostyrko, Gerard Pons-Moll arxiv

Generating realistic and controllable 3D human avatars is a long-standing challenge, particularly when covering broad attribute ranges such as ethnicity, age, clothing styles, and detailed body shapes. Capturing and annotating large-scale human datasets for training generative models is prohibitively expensive and limited in scale and diversity. The central question we address in this paper is: Can existing foundation models be distilled to generate theoretically unbounded, richly annotated 3D human data? We introduce InfiniHuman, a framework that synergistically distills these models to produce richly annotated human data at minimal cost and with theoretically unlimited scalability. We propose InfiniHumanData, a fully automatic pipeline that leverages vision-language and image generation models to create a large-scale multi-modal dataset. User study shows our automatically generated identities are undistinguishable from scan renderings. InfiniHumanData contains 111K identities spanning unprecedented diversity. Each identity is annotated with multi-granularity text descriptions, multi-view RGB images, detailed clothing images, and SMPL body-shape parameters. Building on this dataset, we propose InfiniHumanGen, a diffusion-based generative pipeline conditioned on text, body shape, and clothing assets. InfiniHumanGen enables fast, realistic, and precisely controllable avatar generation. Extensive experiments demonstrate significant improvements over state-of-the-art methods in visual quality, generation speed, and controllability. Our approach enables high-quality avatar generation with fine-grained control at effectively unbounded scale through a practical and affordable solution. We will publicly release the automatic data generation pipeline, the comprehensive InfiniHumanData dataset, and the InfiniHumanGen models at https://yuxuan-xue.com/infini-human.

📄 PDF Abstract BibTeX arXiv:2510.11650

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Proc-GS: Procedural Building Generation for City Assembly with 3D Gaussians

2024-12-10 · Yixuan Li, Xingjian Ran, Linning Xu, Tao Lu 외

Buildings are primary components of cities, often featuring repeated elements such as windows and doors. Traditional 3D building asset creation is labor-intensive and requires specialized skills to develop design rules. …

Asset ManagementManagement

Infinite Motion: Extended Motion Generation via Long Text Instructions

2024-07-11 · Mengtian Li, Chengshuo Zhai, Shengxiang Yao, Zhifeng Xie 외

In the realm of motion generation, the creation of long-duration, high-quality motion sequences remains a significant challenge. This paper presents our groundbreaking work on "Infinite Motion", a novel approach that lev…

Motion GenerationMotion Synthesis

MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation

2026-05-27 · Muyao Wang, Zeke Xie, Yanhao Chen, Lixin Xiu 외 arxiv

End-to-end manga generation is a structured visual storytelling task that requires story decomposition, recurring character and scene grounding, page layout design, panel rendering, page composition, and lettering. Howev…

Visual Storytelling

MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice

2025-03-07 · Hongwei Yi, Tian Ye, Shitong Shao, Xuancheng Yang 외

We present MagicInfinite, a novel diffusion Transformer (DiT) framework that overcomes traditional portrait animation limitations, delivering high-fidelity results across diverse character types-realistic humans, full-bo…

DenoisingPortrait AnimationVideo Generation

MOSAAIC: Managing Optimization towards Shared Autonomy, Authority, and Initiative in Co-creation

2025-05-16 · Alayt Issak, Jeba Rezwana, Casper Harteveld

Striking the appropriate balance between humans and co-creative AI is an open research question in computational creativity. Co-creativity, a form of hybrid intelligence where both humans and AI take action proactively, …

Systematic Literature Review