paper-with-me

Papers

FashionFlow: Leveraging Diffusion Models for Dynamic Fashion Video Synthesis from Static Imagery

2023-09-29 · Tasin Islam, Alina Miron, Xiaohui Liu, Yongmin Li

Our study introduces a new image-to-video generator called FashionFlow to generate fashion videos. By utilising a diffusion model, we are able to create short videos from still fashion images. Our approach involves developing and connecting relevant components with the diffusion model, which results in the creation of high-fidelity videos that are aligned with the conditional image. The components include the use of pseudo-3D convolutional layers to generate videos efficiently. VAE and CLIP encoders capture vital characteristics from still images to condition the diffusion model at a global level. Our research demonstrates a successful synthesis of fashion videos featuring models posing from various angles, showcasing the fit and appearance of the garment. Our findings hold great promise for improving and enhancing the shopping experience for the online fashion industry.

📄 PDF Abstract BibTeX arXiv:2310.00106

Code (1)

1702609/fashionflow 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DreamPose: Fashion Image-to-Video Synthesis via Stable Diffusion

2023-04-12 · Johanna Karras, Aleksander Holynski, Ting-Chun Wang, Ira Kemelmacher-Shlizerman

We present DreamPose, a diffusion-based method for generating animated fashion videos from still images. Given an image and a sequence of human body poses, our method synthesizes a video containing both human and fabric …

DreamPose: Fashion Video Synthesis with Stable Diffusion

2023-01-01 · ICCV 2023 1 · Johanna Karras, Aleksander Holynski, Ting-Chun Wang, Ira Kemelmacher-Shlizerman

We present DreamPose, a diffusion-based method for generating animated fashion videos from still images. Given an image and a sequence of human body poses, our method synthesizes a video containing both human and fab…

ProFashion: Prototype-guided Fashion Video Generation with Multiple Reference Images

2025-05-10 · Xianghao Kong, Qiaosong Qi, Yuanbin Wang, Anyi Rao 외

Fashion video generation aims to synthesize temporally consistent videos from reference images of a designated character. Despite significant progress, existing diffusion-based methods only support a single reference ima…

DenoisingVideo Generation

Fashion-VDM: Video Diffusion Model for Virtual Try-On

2024-10-31 · Johanna Karras, Yingwei Li, Nan Liu, Luyang Zhu 외

We present Fashion-VDM, a video diffusion model (VDM) for generating virtual try-on videos. Given an input garment image and person video, our method aims to generate a high-quality try-on video of the person wearing the…

Video GenerationVirtual Try-on

NVS-Solver: Video Diffusion Model as Zero-Shot Novel View Synthesizer

2024-05-24 · Meng You, Zhiyu Zhu, Hui Liu, Junhui Hou

By harnessing the potent generative capabilities of pre-trained large video diffusion models, we propose NVS-Solver, a new novel view synthesis (NVS) paradigm that operates \textit{without} the need for training. NVS-Sol…

Novel View Synthesis