paper-with-me

홈 › Papers

PromptTea: Let Prompts Tell TeaCache the Optimal Threshold

2025-07-09 · Zishen Huang, Chunyu Yang, Mengyuan Ren arxiv

Despite recent progress in video generation, inference speed remains a major bottleneck. A common acceleration strategy involves reusing model outputs via caching mechanisms at fixed intervals. However, we find that such fixed-frequency reuse significantly degrades quality in complex scenes, while manually tuning reuse thresholds is inefficient and lacks robustness. To address this, we propose Prompt-Complexity-Aware (PCA) caching, a method that automatically adjusts reuse thresholds based on scene complexity estimated directly from the input prompt. By incorporating prompt-derived semantic cues, PCA enables more adaptive and informed reuse decisions than conventional caching methods. We also revisit the assumptions behind TeaCache and identify a key limitation: it suffers from poor input-output relationship modeling due to an oversimplified prior. To overcome this, we decouple the noisy input, enhance the contribution of meaningful textual information, and improve the model's predictive accuracy through multivariate polynomial feature expansion. To further reduce computational cost, we replace the static CFGCache with DynCFGCache, a dynamic mechanism that selectively reuses classifier-free guidance (CFG) outputs based on estimated output variations. This allows for more flexible reuse without compromising output quality. Extensive experiments demonstrate that our approach achieves significant acceleration-for example, 2.79x speedup on the Wan2.1 model-while maintaining high visual fidelity across a range of scenes.

📄 PDF Abstract BibTeX arXiv:2507.06739

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

ACID: Adaptive Caching for vIDeo generation

2026-07-14 · Om Agrawal, Saurabh Agarwal, Aditya Akella arxiv

Video diffusion models produce high-quality generations but remain slow at inference due to their sequential denoising procedure. Caching-based acceleration methods address this by reusing intermediate model outputs: lea…

Video Generation

Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model

2024-11-28 · CVPR 2025 1 · Feng Liu, Shiwei Zhang, XiaoFeng Wang, Yujie Wei 외

As a fundamental backbone for video generation, diffusion models are challenged by low inference speed due to the sequential nature of denoising. Previous methods speed up the models by caching and reusing model outputs …

DenoisingVideo Generation

A Type II Fuzzy Entropy Based Multi-Level Image Thresholding Using Adaptive Plant Propagation Algorithm

2017-08-23 · Sayan Nag

One of the most straightforward, direct and efficient approaches to Image Segmentation is Image Thresholding. Multi-level Image Thresholding is an essential viewpoint in many image processing and Pattern Recognition base…

Image SegmentationSemantic Segmentation

Asynchronous Parallel Empirical Variance Guided Algorithms for the Thresholding Bandit Problem

2017-04-15 · Jie Zhong, Yijun Huang, Ji Liu

This paper considers the multi-armed thresholding bandit problem -- identifying all arms whose expected rewards are above a predefined threshold via as few pulls (or rounds) as possible -- proposed by Locatelli et al. [2…

Super-resolution Probabilistic Rain Prediction from Satellite Data Using 3D U-Nets and EarthFormers

2022-12-06 · Yang Li, Haiyu Dong, Zuliang Fang, Jonathan Weyn 외

Accurate and timely rain prediction is crucial for decision making and is also a challenging task. This paper presents a solution which won the 2 nd prize in the Weather4cast 2022 NeurIPS competition using 3D U-Nets and …

Decision MakingPredictionSuper-Resolution