paper-with-me

Papers

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans

2026-07-30 · Renlong Wu, Haoran Chen, Yuxiang Wei, Xiaowei Jin, Wangmeng Zuo, Hui Li arxiv

Generating high-quality 360-degree dynamic human assets from text prompts is challenging. Existing methods usually synthesize monocular or multi-view videos first and then fit a 4D representation, which is expensive and often causes incomplete geometry or view-inconsistent renderings. We present 4DHumanDiff, a diffusion framework that directly generates dynamic humans represented by 4D Gaussian Splatting (4DGS) from text prompts. By modeling the structured 4D representation space end-to-end, 4DHumanDiff avoids video pre-generation and per-scene reconstruction, making it better suited for view-consistent and temporally coherent asset generation. The model uses a 3D U-Net backbone with temporal attention for motion-aware generation. We further construct a large-scale text-to-4DGS dataset with 60,000 high-quality pairs, and introduce 2D regularization and training-free 4D interpolation to improve rendering quality and motion smoothness. Experiments show that 4DHumanDiff generates consistent 360-degree dynamic humans within one minute, achieves better temporal and multi-view consistency, and reduces inference time by more than 10x.

📄 PDF Abstract BibTeX arXiv:2607.27634

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation

2023-09-07 · Zhuqiang Lu, Kun Hu, Chaoyue Wang, Lei Bai 외

A 360-degree (omni-directional) image provides an all-encompassing spherical view of a scene. Recently, there has been an increasing interest in synthesising 360-degree images from conventional narrow field of view (NFoV…

Image Generation

A Survey on Text-Driven 360-Degree Panorama Generation

2025-02-20 · Hai Wang, Xiaoyu Xiang, Weihao Xia, Jing-Hao Xue

The advent of text-driven 360-degree panorama generation, enabling the synthesis of 360-degree panoramic images directly from textual descriptions, marks a transformative advancement in immersive visual content creation.…

Scene GenerationSurvey

PairHuman: A High-Fidelity Photographic Dataset for Customized Dual-Person Generation

2025-11-20 · Ting Pan, Ye Wang, Peiguang Jing, Rui Ma 외 arxiv

Personalized dual-person portrait customization has considerable potential applications, such as preserving emotional memories and facilitating wedding photography planning. However, the absence of a benchmark dataset hi…

Pose-variant 3D Facial Attribute Generation

2019-07-24 · Feng-Ju Chang, Xiang Yu, Ram Nevatia, Manmohan Chandraker

We address the challenging problem of generating facial attributes using a single image in an unconstrained pose. In contrast to prior works that largely consider generation on 2D near-frontal images, we propose a GAN-ba…

3D ReconstructionAttributeGenerative Adversarial Network

OPa-Ma: Text Guided Mamba for 360-degree Image Out-painting

2024-07-15 · Penglei Gao, Kai Yao, Tiandi Ye, Steven Wang 외

In this paper, we tackle the recently popular topic of generating 360-degree images given the conventional narrow field of view (NFoV) images that could be taken from a single camera or cellphone. This task aims to predi…

Image GenerationMamba