paper-with-me

Papers

SpinMeRound: Consistent Multi-View Identity Generation Using Diffusion Models

2025-04-14 · Stathis Galanakis, Alexandros Lattas, Stylianos Moschoglou, Bernhard Kainz, Stefanos Zafeiriou

Despite recent progress in diffusion models, generating realistic head portraits from novel viewpoints remains a significant challenge. Most current approaches are constrained to limited angular ranges, predominantly focusing on frontal or near-frontal views. Moreover, although the recent emerging large-scale diffusion models have been proven robust in handling 3D scenes, they underperform on facial data, given their complex structure and the uncanny valley pitfalls. In this paper, we propose SpinMeRound, a diffusion-based approach designed to generate consistent and accurate head portraits from novel viewpoints. By leveraging a number of input views alongside an identity embedding, our method effectively synthesizes diverse viewpoints of a subject whilst robustly maintaining its unique identity features. Through experimentation, we showcase our model's generation capabilities in 360 head synthesis, while beating current state-of-the-art multiview diffusion models.

📄 PDF Abstract BibTeX arXiv:2504.10716

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

HarmoView: Harmonizing Multi-View Constraints for Identity-Consistent Video Generation

2026-06-09 · Cong Wang, Zhentao Yu, Hongmei Wang, Weicong Liang 외 arxiv

Current identity-consistent video generation methods struggle to preserve appearance fidelity under large viewpoint changes. While introducing multi-view reference input offers a natural solution, progress remains constr…

Spatial ReasoningVideo Generation

MagicView: Multi-View Consistent Identity Customization via Priors-Guided In-Context Learning

2025-10-31 · Hengjia Li, Jianjin Xu, Keli Cheng, Lei Wang 외 arxiv

Recent advances in personalized generative models have demonstrated impressive capabilities in producing identity-consistent images of the same individual across diverse scenes. However, most existing methods lack explic…

Semantic correspondence

HumanOrbit: 3D Human Reconstruction as 360° Orbit Generation

2026-02-27 · Keito Suzuki, Kunyao Chen, Lei Wang, Bang Du 외 arxiv

We present a method for generating a full 360° orbit video around a person from a single input image. Existing methods typically adapt image-based diffusion models for multi-view synthesis, but yield inconsistent results…

3D Human ReconstructionImage Generation

WildActor: Unconstrained Identity-Preserving Video Generation

2026-02-28 · Qin Guo, Tianyu Yang, Xuanhua He, Fei Shen 외 arxiv

Production-ready human video generation requires digital actors to maintain strictly consistent full-body identities across dynamic shots, viewpoints and motions, a setting that remains challenging for existing methods. …

Video Generation

ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation

2026-02-10 · Mingyang Wu, Ashirbad Mishra, Soumik Dey, Shuo Xing 외 arxiv

Image-to-Video generation (I2V) animates a static image into a temporally coherent video sequence following textual instructions, yet preserving fine-grained object identity under changing viewpoints remains a persistent…

Video Generation