paper-with-me

Papers

Versatile Transition Generation with Image-to-Video Diffusion

2025-08-03 · Zuhao Yang, Jiahui Zhang, Yingchen Yu, Shijian Lu, Song Bai arxiv

Leveraging text, images, structure maps, or motion trajectories as conditional guidance, diffusion models have achieved great success in automated and high-quality video generation. However, generating smooth and rational transition videos given the first and last video frames as well as descriptive text prompts is far underexplored. We present VTG, a Versatile Transition video Generation framework that can generate smooth, high-fidelity, and semantically coherent video transitions. VTG introduces interpolation-based initialization that helps preserve object identity and handle abrupt content changes effectively. In addition, it incorporates dual-directional motion fine-tuning and representation alignment regularization to mitigate the limitations of pre-trained image-to-video diffusion models in motion smoothness and generation fidelity, respectively. To evaluate VTG and facilitate future studies on unified transition generation, we collected TransitBench, a comprehensive benchmark for transition generation covering two representative transition tasks: concept blending and scene transition. Extensive experiments show that VTG achieves superior transition performance consistently across all four tasks.

📄 PDF Abstract BibTeX arXiv:2508.01698

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control

2025-01-07 · Zekai Gu, Rui Yan, Jiahao Lu, Peng Li 외

Diffusion models have demonstrated impressive performance in generating high-quality videos from text prompts or images. However, precise control over the video generation process, such as camera manipulation or content …

Video Generation

TVG: A Training-free Transition Video Generation Method with Diffusion Models

2024-08-24 · Rui Zhang, Yaosen Chen, Yuegen Liu, Wei Wang 외

Transition videos play a crucial role in media production, enhancing the flow and coherence of visual narratives. Traditional methods like morphing often lack artistic appeal and require specialized skills, limiting thei…

GPRVideo Generation

ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer

2026-03-16 · Ruonan Yu, Zhenxiong Tan, Zigeng Chen, Songhua Liu 외 arxiv

Diffusion Transformers (DiTs) have demonstrated remarkable scalability and quality in image and video generation, prompting growing interest in extending them to controllable generation and editing tasks. However, compar…

Video Generation

SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction

2023-10-31 · Xinyuan Chen, Yaohui Wang, Lingjun Zhang, Shaobin Zhuang 외

Recently video generation has achieved substantial progress with realistic results. Nevertheless, existing AI-generated videos are usually very short clips ("shot-level") depicting a single scene. To deliver a coherent l…

PredictionSemantic SimilaritySemantic Textual SimilarityVideo Generation+1

Uniform Discrete Diffusion with Metric Path for Video Generation

2025-10-28 · Haoge Deng, Ting Pan, Fan Zhang, Yang Liu 외 arxiv

Continuous-space video generation has advanced rapidly, while discrete approaches lag behind due to error accumulation and long-context inconsistency. In this work, we revisit discrete generative modeling and present Uni…

Video GenerationImage Generation