paper-with-me

Papers

Paris 2.0: A Decentralized Diffusion Model for Video Generation

2026-05-25 · Ali Rouzbayani, Bidhan Roy, Marcos Villagra, Zhiying Jiang arxiv

We present Paris 2.0, the first video generation model pre-trained through decentralized computation. Its training recipe builds upon Paris 1.0 (arXiv:2510.03434), the first ever open-weight Decentralized Diffusion Model (DDM), which showed that image generation can be trained without a monolithic GPU cluster. However, temporally coherent video generation had remained an open problem under decentralized training, and Paris 2.0 closes it. In low-resolution text-to-video training, against a monolithic model trained on the same data under a matched total compute budget, Paris 2.0 cuts Frechet Video Distance (FVD) from 561.04 to 279.01, a ~2.0x improvement, and lifts CLIP text-video similarity and aesthetic score.

📄 PDF Abstract BibTeX arXiv:2605.26064

Code (0)

등록된 구현이 없습니다.

Tasks

Video GenerationImage Generation

Similar Papers 제목 키워드 기반

Paris: A Decentralized Trained Open-Weight Diffusion Model

2025-10-03 · Zhiying Jiang, Raihan Seraj, Marcos Villagra, Bidhan Roy arxiv

We present Paris, the first publicly released diffusion model pre-trained entirely through decentralized computation. Paris demonstrates that high-quality text-to-image generation can be achieved without centrally coordi…

Text-to-Image Generation

From Image to Video: An Empirical Study of Diffusion Representations

2025-02-10 · Pedro Vélez, Luisa F. Polanía, Yi Yang, Chuhan Zhang 외

Diffusion models have revolutionized generative modeling, enabling unprecedented realism in image and video synthesis. This success has sparked interest in leveraging their representations for visual understanding tasks.…

Action RecognitionDepth Estimationimage-classificationImage Classification+2

FrameBridge: Improving Image-to-Video Generation with Bridge Models

2024-10-20 · Yuji Wang, Zehua Chen, Xiaoyu Chen, Jun Zhu 외

Image-to-video (I2V) generation is gaining increasing attention with its wide application in video synthesis. Recently, diffusion-based I2V models have achieved remarkable progress given their novel design on network arc…

Image AnimationImage to Video GenerationVideo Generation

Discriminator-Free Direct Preference Optimization for Video Diffusion

2025-04-11 · Haoran Cheng, Qide Dong, Liang Peng, Zhizhou Sha 외

Direct Preference Optimization (DPO), which aligns models with human preferences through win/lose data pairs, has achieved remarkable success in language and image generation. However, applying DPO to video diffusion mod…

Image Generation

4Dynamic: Text-to-4D Generation with Hybrid Priors

2024-07-17 · Yu-Jie Yuan, Leif Kobbelt, Jiwen Liu, Yuan Zhang 외

Due to the fascinating generative performance of text-to-image diffusion models, growing text-to-3D generation works explore distilling the 2D generative priors into 3D, using the score distillation sampling (SDS) loss, …

3D GenerationText to 3D