paper-with-me

홈 › Papers

From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition

2025-09-24 · Ling Lo, Kelvin C. K. Chan, Wen-Huang Cheng, Ming-Hsuan Yang arxiv

Existing models often struggle with complex temporal changes, particularly when generating videos with gradual attribute transitions. The most common prompt interpolation approach for motion transitions often fails to handle gradual attribute transitions, where inconsistencies tend to become more pronounced. In this work, we propose a simple yet effective method to extend existing models for smooth and consistent attribute transitions, through introducing frame-wise guidance during the denoising process. Our approach constructs a data-specific transitional direction for each noisy latent, guiding the gradual shift from initial to final attributes frame by frame while preserving the motion dynamics of the video. Moreover, we present the Controlled-Attribute-Transition Benchmark (CAT-Bench), which integrates both attribute and motion dynamics, to comprehensively evaluate the performance of different models. We further propose two metrics to assess the accuracy and smoothness of attribute transitions. Experimental results demonstrate that our approach performs favorably against existing baselines, achieving visual fidelity, maintaining alignment with text prompts, and delivering seamless attribute transitions. Code and CATBench are released: https://github.com/lynn-ling-lo/Prompt2Progression.

📄 PDF Abstract BibTeX arXiv:2509.19690

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation

2025-04-30 · Haiyang Zhou, Wangbo Yu, Jiawen Guan, Xinhua Cheng 외

The rapid advancement of diffusion models holds the promise of revolutionizing the application of VR and AR technologies, which typically require scene-level 4D assets for user experience. Nonetheless, existing diffusion…

Depth EstimationScene GenerationVideo Generation

Medical Video Generation for Disease Progression Simulation

2024-11-18 · Xu Cao, Kaizhao Liang, Kuei-Da Liao, Tianren Gao 외

Modeling disease progression is crucial for improving the quality and efficacy of clinical diagnosis and prognosis, but it is often hindered by a lack of longitudinal medical image monitoring for individual patients. To …

PrognosisVideo Generation

Prompt-Adapter Context Routing for Parameter-Efficient Multi-Shot Long Video Extrapolation

2026-07-07 · Anna Córdoba, Adam Puente Tercero, Nerea Angulo Hijo, Mar Linares Tercero 외 arxiv

We present PACR-Video, a parameter-efficient framework for multi-shot long video extrapolation that preserves recurring entities, scene structure, visual style, and causal progression without full generator fine-tuning. …

ChronosObserver: Taming 4D World with Hyperspace Diffusion Sampling

2025-12-01 · Qisen Wang, Yifan Zhao, Peisen Shen, Jialu Li 외 arxiv

Although prevailing camera-controlled video generation models can produce cinematic results, lifting them directly to the generation of 3D-consistent and high-fidelity time-synchronized multi-view videos remains challeng…

Data AugmentationVideo Generation

FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis

2023-12-29 · CVPR 2024 1 · Feng Liang, Bichen Wu, Jialiang Wang, Licheng Yu 외

Diffusion models have transformed the image-to-image (I2I) synthesis and are now permeating into videos. However, the advancement of video-to-video (V2V) synthesis has been hampered by the challenge of maintaining tempor…

Optical Flow EstimationVideo-to-Video Synthesis