paper-with-me

Papers

Multi-View Data Generation Without View Supervision

2017-11-01 · ICLR 2018 1 · Mickaël Chen, Ludovic Denoyer, Thierry Artières

The development of high-dimensional generative models has recently gained a great surge of interest with the introduction of variational auto-encoders and generative adversarial neural networks. Different variants have been proposed where the underlying latent space is structured, for example, based on attributes describing the data to generate. We focus on a particular problem where one aims at generating samples corresponding to a number of objects under various views. We assume that the distribution of the data is driven by two independent latent factors: the content, which represents the intrinsic features of an object, and the view, which stands for the settings of a particular observation of that object. Therefore, we propose a generative model and a conditional variant built on such a disentangled latent space. This approach allows us to generate realistic samples corresponding to various objects in a high variety of views. Unlike many multi-view approaches, our model doesn't need any supervision on the views but only on the content. Compared to other conditional generation approaches that are mostly based on binary or categorical attributes, we make no such assumption about the factors of variations. Our model can be used on problems with a huge, potentially infinite, number of categories. We experiment it on four image datasets on which we demonstrate the effectiveness of the model and its ability to generalize.

📄 PDF Abstract BibTeX arXiv:1711.00305

Code (1)

mickaelChen/GMV 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation

2024-04-26 · SeungWook Kim, Yichun Shi, Kejie Li, Minsu Cho 외

Using image as prompts for 3D generation demonstrate particularly strong performances compared to using text prompts alone, for images provide a more intuitive guidance for the 3D generation process. In this work, we del…

3D Generation

Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation

2026-05-24 · Aviral Chharia, Fernando De la Torre arxiv

High-fidelity 3D Gaussian head avatar generation is critical for applications such as AR/VR, telepresence, and digital humans. Existing methods depend on multi-view datasets, 3D captures, or intermediate 2D view synthesi…

ViewFusion: Towards Multi-View Consistency via Interpolated Denoising

2024-02-29 · CVPR 2024 1 · Xianghui Yang, Yan Zuo, Sameera Ramasinghe, Loris Bazzani 외

Novel-view synthesis through diffusion models has demonstrated remarkable potential for generating diverse and high-quality images. Yet, the independent process of image generation in these prevailing methods leads to ch…

DenoisingImage GenerationNovel View Synthesis

SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency

2024-07-24 · Yiming Xie, Chun-Han Yao, Vikram Voleti, Huaizu Jiang 외

We present Stable Video 4D (SV4D), a latent video diffusion model for multi-frame and multi-view consistent dynamic 3D content generation. Unlike previous methods that rely on separately trained generative models for vid…

NeRFNovel View SynthesisVideo Generation

StreetDiff: Multi-view Street Scenes Generation via Cross-view Consistent Multi-view Stable Diffusion with Structure Prompts

2026-09-09 · Qi Zhang, Yanyifan Wang, Weiyuan Zhang, Hui Huang arxiv

Multi-view diffusion models have shown strong performance in scenes with strong geometric priors and sparse semantics, such as indoor rooms or simple outdoor environments (e.g., fields, courtyards). However, they often f…

Scene Generation