paper-with-me

홈 › Papers

HuGDiffusion: Generalizable Single-Image Human Rendering via 3D Gaussian Diffusion

2025-01-25 · Yingzhi Tang, Qijian Zhang, Junhui Hou

We present HuGDiffusion, a generalizable 3D Gaussian splatting (3DGS) learning pipeline to achieve novel view synthesis (NVS) of human characters from single-view input images. Existing approaches typically require monocular videos or calibrated multi-view images as inputs, whose applicability could be weakened in real-world scenarios with arbitrary and/or unknown camera poses. In this paper, we aim to generate the set of 3DGS attributes via a diffusion-based framework conditioned on human priors extracted from a single image. Specifically, we begin with carefully integrated human-centric feature extraction procedures to deduce informative conditioning signals. Based on our empirical observations that jointly learning the whole 3DGS attributes is challenging to optimize, we design a multi-stage generation strategy to obtain different types of 3DGS attributes. To facilitate the training process, we investigate constructing proxy ground-truth 3D Gaussian attributes as high-quality attribute-level supervision signals. Through extensive experiments, our HuGDiffusion shows significant performance improvements over the state-of-the-art methods. Our code will be made publicly available.

📄 PDF Abstract BibTeX arXiv:2501.15008

Code (0)

등록된 구현이 없습니다.

Tasks

3DGSAttributeNovel View Synthesis

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

ConTex-Human: Free-View Rendering of Human from a Single Image with Texture-Consistent Synthesis

2023-11-28 · CVPR 2024 1 · Xiangjun Gao, Xiaoyu Li, Chaopeng Zhang, Qi Zhang 외

In this work, we propose a method to address the challenge of rendering a 3D human from a single image in a free-view manner. Some existing approaches could achieve this by using generalizable pixel-aligned implicit fiel…

Generalizable Neural Human Renderer

2024-04-22 · Mana Masuda, Jinhyung Park, Shun Iwase, Rawal Khirodkar 외

While recent advancements in animatable human rendering have achieved remarkable results, they require test-time optimization for each subject which can be a significant limitation for real-world applications. To address…

EG-HumanNeRF: Efficient Generalizable Human NeRF Utilizing Human Prior for Sparse View

2024-10-16 · Zhaorong Wang, Yoshihiro Kanamori, Yuki Endo

Generalizable neural radiance field (NeRF) enables neural-based digital human rendering without per-scene retraining. When combined with human prior knowledge, high-quality human rendering can be achieved even with spars…

NeRFNovel View Synthesis

LIFe-GoM: Generalizable Human Rendering with Learned Iterative Feedback Over Multi-Resolution Gaussians-on-Mesh

2025-02-13 · Jing Wen, Alexander G. Schwing, Shenlong Wang

Generalizable rendering of an animatable human avatar from sparse inputs relies on data priors and inductive biases extracted from training on large data to avoid scene-specific optimization and to enable fast reconstruc…

Global Latent Neural Rendering

2023-12-13 · CVPR 2024 1 · Thomas Tanay, Matteo Maggioni

A recent trend among generalizable novel view synthesis methods is to learn a rendering operator acting over single camera rays. This approach is promising because it removes the need for explicit volumetric rendering, b…

Generalizable Novel View SynthesisNeural RenderingNovel View Synthesis