paper-with-me

홈 › Papers

One Transform To Compute Them All: Efficient Fusion-Based Full-Reference Video Quality Assessment

2023-04-06 · Abhinau K. Venkataramanan, Cosmin Stejerean, Ioannis Katsavounidis, Alan C. Bovik

The Visual Multimethod Assessment Fusion (VMAF) algorithm has recently emerged as a state-of-the-art approach to video quality prediction, that now pervades the streaming and social media industry. However, since VMAF requires the evaluation of a heterogeneous set of quality models, it is computationally expensive. Given other advances in hardware-accelerated encoding, quality assessment is emerging as a significant bottleneck in video compression pipelines. Towards alleviating this burden, we propose a novel Fusion of Unified Quality Evaluators (FUNQUE) framework, by enabling computation sharing and by using a transform that is sensitive to visual perception to boost accuracy. Further, we expand the FUNQUE framework to define a collection of improved low-complexity fused-feature models that advance the state-of-the-art of video quality performance with respect to both accuracy, by 4.2\% to 5.3\%, and computational efficiency, by factors of 3.8 to 11 times!

📄 PDF Abstract BibTeX arXiv:2304.03412

Code (0)

등록된 구현이 없습니다.

Tasks

AllComputational EfficiencyVideo CompressionVideo Quality Assessment

Similar Papers 제목 키워드 기반

Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis

2024-09-04 · Aishwarya Agarwal, Srikrishna Karanam, Balaji Vasan Srinivasan

We consider the problem of independently, in a disentangled fashion, controlling the outputs of text-to-image diffusion models with color and style attributes of a user-supplied reference image. We present the first trai…

DisentanglementImage Generation

Anchoring Instruction Outside Mask: Exact Reference Caching for Efficient In-Context Diffusion Transformers

2026-08-21 · Yangshuai Liu, Zheming Li, Jiaao Li, Kang He 외 arxiv

Omnimodal generation is central to a wide range of content creation and editing applications. In-context conditioning is essential to this paradigm. It allows diffusion transformers to process text instructions and visua…

Instruction Following

SoftCap: Soft-Budget Control for Diffusion Transformer Acceleration

2026-05-26 · Yuhang Zhang, Junxiang Qiu, Huixia Ben, Zhenhua Tang 외 arxiv

Diffusion Transformers (DiTs) achieve strong visual quality, but their iterative denoising process requires many costly Transformer evaluations. Training-free acceleration methods reduce this cost by caching, forecasting…

Zero-shot Segmentation of Skin Conditions: Erythema with Edit-Friendly Inversion

2025-08-02 · Konstantinos Moutselos, Ilias Maglogiannis arxiv

This study proposes a zero-shot image segmentation framework for detecting erythema (redness of the skin) using edit-friendly inversion in diffusion models. The method synthesizes reference images of the same patient tha…

Image Segmentation

Controllable Texture Tiling with Transformed RoPE-Enhanced Diffusion Models

2026-06-22 · Junrong Huang, Zhiyuan Zhang, Rui Tang, Hongbo Fu 외 arxiv

Realistic integration of user-specified textures into scene images is a fundamental task in computer graphics and image editing. While existing material transfer and reference-guided inpainting methods can edit surface a…

Image Editing