paper-with-me

Papers

A Preprocessing Framework for Video Machine Vision under Compression

2025-12-17 · Fei Zhao, Mengxi Guo, Shijie Zhao, Junlin Li, Li Zhang, Xiaodong Xie arxiv

There has been a growing trend in compressing and transmitting videos from terminals for machine vision tasks. Nevertheless, most video coding optimization method focus on minimizing distortion according to human perceptual metrics, overlooking the heightened demands posed by machine vision systems. In this paper, we propose a video preprocessing framework tailored for machine vision tasks to address this challenge. The proposed method incorporates a neural preprocessor which retaining crucial information for subsequent tasks, resulting in the boosting of rate-accuracy performance. We further introduce a differentiable virtual codec to provide constraints on rate and distortion during the training stage. We directly apply widely used standard codecs for testing. Therefore, our solution can be easily applied to real-world scenarios. We conducted extensive experiments evaluating our compression method on two typical downstream tasks with various backbone networks. The experimental results indicate that our approach can save over 15% of bitrate compared to using only the standard codec anchor version.

📄 PDF Abstract BibTeX arXiv:2512.15331

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Tri-Dynamic Preprocessing Framework for UGC Video Compression

2025-12-18 · Fei Zhao, Mengxi Guo, Shijie Zhao, Junlin Li 외 arxiv

In recent years, user generated content (UGC) has become the dominant force in internet traffic. However, UGC videos exhibit a higher degree of variability and diverse characteristics compared to traditional encoding tes…

Rate-Perception Optimized Preprocessing for Video Coding

2023-01-25 · Chengqian Ma, Zhiqiang Wu, Chunlei Cai, Pengwei Zhang 외

In the past decades, lots of progress have been done in the video compression field including traditional video codec and learning-based video codec. However, few studies focus on using preprocessing techniques to improv…

Image Quality AssessmentVideo Compression

Preprocessing Enhanced Image Compression for Machine Vision

2022-06-12 · Guo Lu, Xingtong Ge, Tianxiong Zhong, Jing Geng 외

Recently, more and more images are compressed and sent to the back-end devices for the machine analysis tasks~(\textit{e.g.,} object detection) instead of being purely watched by humans. However, most traditional or lear…

Image Compressionobject-detectionObject DetectionQuantization

EasyVideoR1: Easier RL for Video Understanding

2026-04-18 · Chuanyu Qin, Chenxu Yang, Qingyi Si, Naibin Gu 외 arxiv

Reinforcement learning from verifiable rewards (RLVR) has demonstrated remarkable effectiveness in improving the reasoning capabilities of large language models. As models evolve into natively multimodal architectures, e…

Reinforcement Learning

Human Action Recognition using Local Two-Stream Convolution Neural Network Features and Support Vector Machines

2020-02-19 · David Torpey, Turgay Celik

This paper proposes a simple yet effective method for human action recognition in video. The proposed method separately extracts local appearance and motion features using state-of-the-art three-dimensional convolutional…

Action ClassificationAction RecognitionOptical Flow EstimationTemporal Action Localization