paper-with-me

홈 › Papers

Lost in Translation: Modern Neural Networks Still Struggle With Small Realistic Image Transformations

2024-04-10 · Ofir Shifman, Yair Weiss

Deep neural networks that achieve remarkable performance in image classification have previously been shown to be easily fooled by tiny transformations such as a one pixel translation of the input image. In order to address this problem, two approaches have been proposed in recent years. The first approach suggests using huge datasets together with data augmentation in the hope that a highly varied training set will teach the network to learn to be invariant. The second approach suggests using architectural modifications based on sampling theory to deal explicitly with image translations. In this paper, we show that these approaches still fall short in robustly handling 'natural' image translations that simulate a subtle change in camera orientation. Our findings reveal that a mere one-pixel translation can result in a significant change in the predicted image representation for approximately 40% of the test images in state-of-the-art models (e.g. open-CLIP trained on LAION-2B or DINO-v2) , while models that are explicitly constructed to be robust to cyclic translations can still be fooled with 1 pixel realistic (non-cyclic) translations 11% of the time. We present Robust Inference by Crop Selection: a simple method that can be proven to achieve any desired level of consistency, although with a modest tradeoff with the model's accuracy. Importantly, we demonstrate how employing this method reduces the ability to fool state-of-the-art models with a 1 pixel translation to less than 5% while suffering from only a 1% drop in classification accuracy. Additionally, we show that our method can be easy adjusted to deal with circular shifts as well. In such case we achieve 100% robustness to integer shifts with state-of-the-art accuracy, and with no need for any further training.

📄 PDF Abstract BibTeX arXiv:2404.07153

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentationimage-classificationImage ClassificationTranslation

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Machine Translation from Spoken Language to Sign Language using Pre-trained Language Model as Encoder

2020-05-01 · LREC 2020 5 · Taro Miyazaki, Yusuke Morita, Masanori Sano

Sign language is the first language for those who were born deaf or lost their hearing in early childhood, so such individuals require services provided with sign language. To achieve flexible open-domain services with s…

Language ModelingLanguage ModellingMachine TranslationTranslation

Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?

2025-03-23 · Aabid Karim, Abdul Karim, Bhoomika Lohana, Matt Keon 외

Large Language Models (LLMs) have significantly advanced various fields, particularly coding, mathematical reasoning, and logical problem solving. However, a critical question remains: Do these mathematical reasoning abi…

GSM8KMathMathematical Reasoning

Scale-Aware Relay and Scale-Adaptive Loss for Tiny Object Detection in Aerial Images

2025-11-13 · Jinfu Li, Yuqi Huang, Hong Song, Ting Wang 외 arxiv

Recently, despite the remarkable advancements in object detection, modern detectors still struggle to detect tiny objects in aerial images. One key reason is that tiny objects carry limited features that are inevitably d…

Object Detection In Aerial Images

Lost in Interpretation: Predicting Untranslated Terminology in Simultaneous Interpretation

2019-04-01 · NAACL 2019 6 · Nikolai Vogler, Craig Stewart, Graham Neubig

Simultaneous interpretation, the translation of speech from one language to another in real-time, is an inherently difficult and strenuous task. One of the greatest challenges faced by interpreters is the accurate transl…

Translation

Lost in Distillation: A Case Study in Toxicity Modeling

2022-07-01 · NAACL (WOAH) 2022 7 · Alyssa Chvasta, Alyssa Lees, Jeffrey Sorensen, Lucy Vasserman 외

In an era of increasingly large pre-trained language models, knowledge distillation is a powerful tool for transferring information from a large model to a smaller one. In particular, distillation is of tremendous benefi…

Knowledge Distillation