paper-with-me

홈 › Papers

RGB no more: Minimally-decoded JPEG Vision Transformers

2022-11-29 · CVPR 2023 1 · Jeongsoo Park, Justin Johnson

Most neural networks for computer vision are designed to infer using RGB images. However, these RGB images are commonly encoded in JPEG before saving to disk; decoding them imposes an unavoidable overhead for RGB networks. Instead, our work focuses on training Vision Transformers (ViT) directly from the encoded features of JPEG. This way, we can avoid most of the decoding overhead, accelerating data load. Existing works have studied this aspect but they focus on CNNs. Due to how these encoded features are structured, CNNs require heavy modification to their architecture to accept such data. Here, we show that this is not the case for ViTs. In addition, we tackle data augmentation directly on these encoded features, which to our knowledge, has not been explored in-depth for training in this setting. With these two improvements -- ViT and data augmentation -- we show that our ViT-Ti model achieves up to 39.2% faster training and 17.9% faster inference with no accuracy loss compared to the RGB counterpart.

📄 PDF Abstract BibTeX arXiv:2211.16421

Code (1)

jeongsoop/rgb-no-more 공식 구현 pytorch

Tasks

Data Augmentation

Similar Papers 제목 키워드 기반

Learning-based Compression for Material and Texture Recognition

2021-04-16 · Yingpeng Deng, Lina J. Karam

Learning-based image compression was shown to achieve a competitive performance with state-of-the-art transform-based codecs. This motivated the development of new learning-based visual compression standards such as JPEG…

domain classificationGeneral ClassificationImage Compression

Embedding Novel Views in a Single JPEG Image

2021-08-30 · ICCV 2021 10 · Yue Wu, Guotao Meng, Qifeng Chen

We propose a novel approach for embedding novel views in a single JPEG image while preserving the perceptual fidelity of the modified JPEG image and the restored novel views. We adopt the popular novel view synthesis rep…

Novel View Synthesis

Forward Error Correction applied to JPEG-XS codestreams

2022-07-11 · Antoine Legrand, Benoît Macq, Christophe De Vleeschouwer

JPEG-XS offers low complexity image compression for applications with constrained but reasonable bit-rate, and low latency. Our paper explores the deployment of JPEG-XS on lossy packet networks. To preserve low latency, …

Image Compression

RAW Image Reconstruction Using a Self-Contained sRGB-JPEG Image With Only 64 KB Overhead

2016-06-01 · CVPR 2016 6 · Rang M. H. Nguyen, Michael S. Brown

Most camera images are saved as 8-bit standard RGB (sRGB) compressed JPEGs. Even when JPEG compression is set to its highest quality, the encoded sRGB image has been significantly processed in terms of color and tone ma…

Image Reconstruction

An End-to-End Compression Framework Based on Convolutional Neural Networks

2017-08-02 · Feng Jiang, Wen Tao, Shaohui Liu, Jie Ren 외

Deep learning, e.g., convolutional neural networks (CNNs), has achieved great success in image processing and computer vision especially in high level vision applications such as recognition and understanding. However, i…

DenoisingImage Compression