paper-with-me

Papers

MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human Reconstruction

2024-12-04 · CVPR 2025 1 · Gangjian Zhang, Nanjie Yao, Shunsi Zhang, HanFeng Zhao, Guoliang Pang, Jian Shu, Hao Wang

This paper investigates the research task of reconstructing the 3D clothed human body from a monocular image. Due to the inherent ambiguity of single-view input, existing approaches leverage pre-trained SMPL(-X) estimation models or generative models to provide auxiliary information for human reconstruction. However, these methods capture only the general human body geometry and overlook specific geometric details, leading to inaccurate skeleton reconstruction, incorrect joint positions, and unclear cloth wrinkles. In response to these issues, we propose a multi-level geometry learning framework. Technically, we design three key components: skeleton-level enhancement, joint-level augmentation, and wrinkle-level refinement modules. Specifically, we effectively integrate the projected 3D Fourier features into a Gaussian reconstruction model, introduce perturbations to improve joint depth estimation during training, and refine the human coarse wrinkles by resembling the de-noising process of diffusion model. Extensive quantitative and qualitative experiments on two out-of-distribution test sets show the superior performance of our approach compared to state-of-the-art (SOTA) methods.

📄 PDF Abstract BibTeX arXiv:2412.03103

Code (0)

등록된 구현이 없습니다.

Tasks

Depth Estimation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

MultiGO++: Monocular 3D Clothed Human Reconstruction via Geometry-Texture Collaboration

2026-03-05 · Nanjie Yao, Gangjian Zhang, Wenhao Shen, Jian Shu 외 arxiv

Monocular 3D clothed human reconstruction aims to generate a complete and realistic textured 3D avatar from a single image. Existing methods are commonly trained under multi-view supervision with annotated geometric prio…

Text2Room: Extracting Textured 3D Meshes from 2D Text-to-Image Models

2023-03-21 · ICCV 2023 1 · Lukas Höllein, Ang Cao, Andrew Owens, Justin Johnson 외

We present Text2Room, a method for generating room-scale textured 3D meshes from a given text prompt as input. To this end, we leverage pre-trained 2D text-to-image models to synthesize a sequence of images from differen…

3D geometryText to 3D

Temporal Consistency Loss for High Resolution Textured and Clothed 3DHuman Reconstruction from Monocular Video

2021-04-19 · Akin Caliskan, Armin Mustafa, Adrian Hilton

We present a novel method to learn temporally consistent 3D reconstruction of clothed people from a monocular video. Recent methods for 3D human reconstruction from monocular video using volumetric, implicit or parametri…

3D geometry3D Human Reconstruction3D Human Shape Estimation3D Reconstruction+1

Monocular 3D Object Reconstruction with GAN Inversion

2022-07-20 · Junzhe Zhang, Daxuan Ren, Zhongang Cai, Chai Kiat Yeo 외

Recovering a textured 3D mesh from a monocular image is highly challenging, particularly for in-the-wild objects that lack 3D ground truths. In this work, we present MeshInversion, a novel framework to improve the recons…

3D Object ReconstructionObjectObject Reconstruction

xCloth: Extracting Template-free Textured 3D Clothes from a Monocular Image

2022-08-27 · Astitva Srivastava, Chandradeep Pokhariya, Sai Sagar Jinka, Avinash Sharma

Existing approaches for 3D garment reconstruction either assume a predefined template for the garment geometry (restricting them to fixed clothing styles) or yield vertex colored meshes (lacking high-frequency textural d…

Garment Reconstruction