paper-with-me

Papers

Hierarchical Patch VAE-GAN: Generating Diverse Videos from a Single Sample

2020-06-22 · NeurIPS 2020 12 · Shir Gur, Sagie Benaim, Lior Wolf

We consider the task of generating diverse and novel videos from a single video sample. Recently, new hierarchical patch-GAN based approaches were proposed for generating diverse images, given only a single sample at training time. Moving to videos, these approaches fail to generate diverse samples, and often collapse into generating samples similar to the training video. We introduce a novel patch-based variational autoencoder (VAE) which allows for a much greater diversity in generation. Using this tool, a new hierarchical video generation scheme is constructed: at coarse scales, our patch-VAE is employed, ensuring samples are of high diversity. Subsequently, at finer scales, a patch-GAN renders the fine details, resulting in high quality videos. Our experiments show that the proposed method produces diverse samples in both the image domain, and the more challenging video domain.

📄 PDF Abstract BibTeX arXiv:2006.12226

Code (3)

shirgur/hp-vae-gan 공식 구현 pytorch
2023-MindSpore-1/ms-code-23 mindspore
SakiRinn/mindspore-hp-vae-gan mindspore

Tasks

DiversityVideo Generation

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Tora: Trajectory-oriented Diffusion Transformer for Video Generation

2024-07-31 · CVPR 2025 1 · Zhenghao Zhang, Junchao Liao, Menghao Li, Zuozhuo Dai 외

Recent advancements in Diffusion Transformer (DiT) have demonstrated remarkable proficiency in producing high-quality video content. Nonetheless, the potential of transformer-based diffusion models for effectively genera…

Video CompressionVideo Generation

FewGAN: Generating from the Joint Distribution of a Few Images

2022-07-18 · Lior Ben-Moshe, Sagie Benaim, Lior Wolf

We introduce FewGAN, a generative model for generating novel, high-quality and diverse images whose patch distribution lies in the joint patch distribution of a small number of N>1 training samples. The method is, in ess…

Quantization

Hierarchical Pyramid Diverse Attention Networks for Face Recognition

2020-06-01 · CVPR 2020 6 · Qiangchang Wang, Tianyi Wu, He Zheng, Guodong Guo

Deep learning has achieved a great success in face recognition (FR), however, few existing models take hierarchical multi-scale local features into consideration. In this work, we propose a hierarchical pyramid diverse a…

Face Recognition

MovieBench: A Hierarchical Movie Level Dataset for Long Video Generation

2024-11-22 · CVPR 2025 1 · Weijia Wu, MingYu Liu, Zeyu Zhu, Xi Xia 외

Recent advancements in video generation models, like Stable Video Diffusion, show promising results, but primarily focus on short, single-scene videos. These models struggle with generating long videos that involve multi…

Video Generation

Human Video Generation from a Single Image with 3D Pose and View Control

2026-02-24 · Tiantian Wang, Chun-Han Yao, Tao Hu, Mallikarjun Byrasandra Ramalinga Reddy 외 arxiv

Recent diffusion methods have made significant progress in generating videos from single images due to their powerful visual generation capabilities. However, challenges persist in image-to-video synthesis, particularly …

Video Generation