paper-with-me

Papers

Optical-Flow Guided Prompt Optimization for Coherent Video Generation

2024-11-23 · CVPR 2025 1 · Hyelin Nam, JaeMin Kim, Dohun Lee, Jong Chul Ye

While text-to-video diffusion models have made significant strides, many still face challenges in generating videos with temporal consistency. Within diffusion frameworks, guidance techniques have proven effective in enhancing output quality during inference; however, applying these methods to video diffusion models introduces additional complexity of handling computations across entire sequences. To address this, we propose a novel framework called MotionPrompt that guides the video generation process via optical flow. Specifically, we train a discriminator to distinguish optical flow between random pairs of frames from real videos and generated ones. Given that prompts can influence the entire video, we optimize learnable token embeddings during reverse sampling steps by using gradients from a trained discriminator applied to random frame pairs. This approach allows our method to generate visually coherent video sequences that closely reflect natural motion dynamics, without compromising the fidelity of the generated content. We demonstrate the effectiveness of our approach across various models.

📄 PDF Abstract BibTeX arXiv:2411.15540

Code (0)

등록된 구현이 없습니다.

Tasks

Optical Flow EstimationVideo Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Flow-Guided Implicit Neural Representation for Motion-Aware Dynamic MRI Reconstruction

2025-11-21 · Baoqing Li, Yuanyuan Liu, Congcong Liu, Qingyong Zhu 외 arxiv

Dynamic magnetic resonance imaging (dMRI) captures temporally-resolved anatomy but is often challenged by limited sampling and motion-induced artifacts. Conventional motion-compensated reconstructions typically rely on p…

MRI Reconstruction

How Do Optical Flow and Textual Prompts Collaborate to Assist in Audio-Visual Semantic Segmentation?

2026-01-13 · Yujian Lee, Peng Gao, Yongqi Xu, Wentao Fan arxiv

Audio-visual semantic segmentation (AVSS) represents an extension of the audio-visual segmentation (AVS) task, necessitating a semantic understanding of audio-visual scenes beyond merely identifying sound-emitting object…

Semantic Segmentation

Successive optimization of optics and post-processing with differentiable coherent PSF operator and field information

2024-12-19 · Zheng Ren, Jingwen Zhou, Wenguan Zhang, Jiapu Yan 외

Recently, the joint design of optical systems and downstream algorithms is showing significant potential. However, existing rays-described methods are limited to optimizing geometric degradation, making it difficult to f…

Temporally coherent completion of dynamic video

2016-11-01 · J.-B. Huang, S. B. Kang, N. Ahuja, J. Kopf

We present an automatic video completion algorithm that synthesizes missing regions in videos in a temporally coherent fashion. Our algorithm can handle dynamic scenes captured using a moving camera. State-of-the-art app…

Optical Flow EstimationVideo Inpainting

LILAC: Language-Conditioned Object-Centric Optical Flow for Open-Loop Trajectory Generation

2026-03-26 · Motonari Kambara, Koki Seno, Tomoya Kaichi, Yanan Wang 외 arxiv

We address language-conditioned robotic manipulation using flow-based trajectory generation, which enables training on human and web videos of object manipulation and requires only minimal embodiment-specific data. This …