paper-with-me

홈 › Papers

Free$^2$Guide: Gradient-Free Path Integral Control for Enhancing Text-to-Video Generation with Large Vision-Language Models

2024-11-26 · JaeMin Kim, Bryan S Kim, Jong Chul Ye

Diffusion models have achieved impressive results in generative tasks like text-to-image (T2I) and text-to-video (T2V) synthesis. However, achieving accurate text alignment in T2V generation remains challenging due to the complex temporal dependency across frames. Existing reinforcement learning (RL)-based approaches to enhance text alignment often require differentiable reward functions or are constrained to limited prompts, hindering their scalability and applicability. In this paper, we propose Free$^2$Guide, a novel gradient-free framework for aligning generated videos with text prompts without requiring additional model training. Leveraging principles from path integral control, Free$^2$Guide approximates guidance for diffusion models using non-differentiable reward functions, thereby enabling the integration of powerful black-box Large Vision-Language Models (LVLMs) as reward model. Additionally, our framework supports the flexible ensembling of multiple reward models, including large-scale image-based models, to synergistically enhance alignment without incurring substantial computational overhead. We demonstrate that Free$^2$Guide significantly improves text alignment across various dimensions and enhances the overall quality of generated videos.

📄 PDF Abstract BibTeX arXiv:2411.17041

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)Text-to-Video GenerationVideo Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

A model-free approach to continuous-time finance

2022-11-28 · Henry Chiu, Rama Cont

We present a non-probabilistic, pathwise approach to continuous-time finance based on causal functional calculus. We introduce a definition of self-financing, free from any integration concept and show that the value of …

model

Adaptive Smoothing Path Integral Control

2020-05-13 · Dominik Thalmeier, Hilbert J. Kappen, Simone Totaro, Vicenç Gómez

In Path Integral control problems a representation of an optimally controlled dynamical system can be formally computed and serve as a guidepost to learn a parametrized policy. The Path Integral Cross-Entropy (PICE) meth…

Equilibrium Propagation Without Limits

2025-11-27 · Elon Litman arxiv

We liberate Equilibrium Propagation (EP) from the limit of infinitesimal perturbations by establishing a finite-nudge foundation for local credit assignment. By modeling network states as Gibbs-Boltzmann distributions ra…

C-Free-Uniform: A Map-Conditioned Trajectory Sampler for Model Predictive Path Integral Control

2025-10-19 · Yukang Cao, Rahul Moorthy, O. Goktug Poyrazoglu, Volkan Isler arxiv

Trajectory sampling is a key component of sampling-based control mechanisms. Trajectory samplers rely on control input samplers, which generate control inputs u from a distribution p(u | x) where x is the current state. …

Vector Field Guided Path Following Control: Singularity Elimination and Global Convergence

2020-03-22

Vector field guided path following (VF-PF) algorithms are fundamental in robot navigation tasks, but may not deliver the desirable performance when robots encounter singular points where the vector field becomes zero. Th…

Robot Navigation