paper-with-me

홈 › Papers

LEHA-CVQAD: Dataset To Enable Generalized Video Quality Assessment of Compression Artifacts

2025-07-05 · Aleksandr Gushchin, Maksim Smirnov, Dmitriy Vatolin, Anastasia Antsiferova arxiv

We propose the LEHA-CVQAD (Large-scale Enriched Human-Annotated Compressed Video Quality Assessment) dataset, which comprises 6,240 clips for compression-oriented video quality assessment. 59 source videos are encoded with 186 codec-preset variants, 1.8M pairwise, and 1.5k MOS ratings are fused into a single quality scale; part of the videos remains hidden for blind evaluation. We also propose Rate-Distortion Alignment Error (RDAE), a novel evaluation metric that quantifies how well VQA models preserve bitrate-quality ordering, directly supporting codec parameter tuning. Testing IQA/VQA methods reveals that popular VQA metrics exhibit high RDAE and lower correlations, underscoring the dataset challenges and utility. The open part and the results of LEHA-CVQAD are available at https://aleksandrgushchin.github.io/lcvqad/

📄 PDF Abstract BibTeX arXiv:2507.03990

Code (0)

등록된 구현이 없습니다.

Tasks

Video Quality Assessment

Similar Papers 제목 키워드 기반

Maximum Likelihood de novo reconstruction of viral populations using paired end sequencing data

2016-04-16

We present MLEHaplo, a maximum likelihood de novo assembly algorithm for reconstructing viral haplotypes in a virus population from paired-end next generation sequencing (NGS) data. Using the pairing information of reads…

AIM 2024 Challenge on Compressed Video Quality Assessment: Methods and Results

2024-08-21 · Maksim Smirnov, Aleksandr Gushchin, Anastasia Antsiferova, Dmitry Vatolin 외

Video quality assessment (VQA) is a crucial task in the development of video compression standards, as it directly impacts the viewer experience. This paper presents the results of the Compressed Video Quality Assessment…

Image ManipulationvalidVideo CompressionVideo Quality Assessment+1

StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video

2026-05-18 · Huajian Zeng, Chaohua Yao, Yuantai Zhang, Jiaqi Yang 외 arxiv

Recovering world space 4D motion of two interacting hands from egocentric video is a fundamental capability for supervising robot policy learning, where wrist trajectories track the end-effector and finger articulations …

AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection

2025-05-26 · Yejin Lee, Joonghyuk Hahn, Hyeseon Ahn, Yo-Sub Han

Implicit hate speech detection is challenging due to its subtlety and reliance on contextual interpretation rather than explicit offensive words. Current approaches rely on contrastive learning, which are shown to be eff…

Contrastive LearningHate Speech Detectionnamed-entity-recognitionNamed Entity Recognition+1

A Space-Time Transformer for Precipitation Nowcasting

2025-11-14 · Levi Harris, Tianlong Chen arxiv

Meteorological agencies around the world rely on real-time flood guidance to issue life-saving advisories and warnings. For decades traditional numerical weather prediction (NWP) models have been state-of-the-art for pre…

Precipitation ForecastingWeather Forecasting