paper-with-me

홈 › Papers

Odo: Depth-Guided Diffusion for Identity-Preserving Body Reshaping

2025-08-18 · Siddharth Khandelwal, Sridhar Kamath, Arjun Jain arxiv

Human shape editing enables controllable transformation of a person's body shape, such as thin, muscular, or overweight, while preserving pose, identity, clothing, and background. Unlike human pose editing, which has advanced rapidly, shape editing remains relatively under-explored. Current approaches typically rely on 3D morphable models or image warping, often introducing unrealistic body proportions, texture distortions, and background inconsistencies due to alignment errors and deformations. A key limitation is the lack of large-scale, publicly available datasets for training and evaluating body shape manipulation methods. In this work, we introduce the first large-scale dataset of 18,573 images across 1523 subjects, specifically designed for controlled human shape editing. It features diverse variations in body shape, including fat, muscular and thin, captured under consistent identity, clothing, and background conditions. Using this dataset, we propose Odo, an end-to-end diffusion-based method that enables realistic and intuitive body reshaping guided by simple semantic attributes. Our approach combines a frozen UNet that preserves fine-grained appearance and background details from the input image with a ControlNet that guides shape transformation using target SMPL depth maps. Extensive experiments demonstrate that our method outperforms prior approaches, achieving per-vertex reconstruction errors as low as 7.5mm, significantly lower than the 13.6mm observed in baseline methods, while producing realistic results that accurately match the desired target shapes.

📄 PDF Abstract BibTeX arXiv:2508.13065

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Visual Persona: Foundation Model for Full-Body Human Customization

2025-03-19 · CVPR 2025 1 · Jisu Nam, Soowon Son, Zhan Xu, Jing Shi 외

We introduce Visual Persona, a foundation model for text-to-image full-body human customization that, given a single in-the-wild human image, generates diverse images of the individual guided by text descriptions. Unlike…

Appearance Transfer

CA-IDD: Cross-Attention Guided Identity-Conditional Diffusion for Identity-Consistent Face Swapping

2026-04-27 · Md Shohel Rana, Tanoy Debnath arxiv

Face swapping aims to optimize realistic facial image generation by leveraging the identity of a source face onto a target face while preserving pose, expression, and context. However, existing methods, especially GAN-ba…

Image GenerationFace Swapping

Diffusion-based Facial Aesthetics Enhancement with 3D Structure Guidance

2025-03-18 · Lisha Li, Jingwen Hou, Weide Liu, Yuming Fang 외

Facial Aesthetics Enhancement (FAE) aims to improve facial attractiveness by adjusting the structure and appearance of a facial image while preserving its identity as much as possible. Most existing methods adopted deep …

Face Model

Diffusion Handles Enabling 3D Edits for Diffusion Models by Lifting Activations to 3D

2024-01-01 · CVPR 2024 1 · Karran Pandey, Paul Guerrero, Matheus Gadelha, Yannick Hold-Geoffroy 외

Diffusion handles is a novel approach to enable 3D object edits on diffusion images requiring only existing pre-trained diffusion models depth estimation without any fine-tuning or 3D object retrieval. The edited res…

3D Object RetrievalDepth EstimationObjectRetrieval

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation

2026-07-22 · Zhiyuan Xia, Haojie Li, Jingyu Lin, Yiguo Qiao 외 arxiv

Thermal-to-visible face translation presents fundamental challenges including geometric discontinuities, semantic attribute mismatches, and identity degradation. We propose MTVDiff, a novel multimodal latent diffusion fr…

Face VerificationFace Recognition