paper-with-me

Papers

CNN Fixations: An unraveling approach to visualize the discriminative image regions

2017-08-22 · Konda Reddy Mopuri, Utsav Garg, R. Venkatesh Babu

Deep convolutional neural networks (CNN) have revolutionized various fields of vision research and have seen unprecedented adoption for multiple tasks such as classification, detection, captioning, etc. However, they offer little transparency into their inner workings and are often treated as black boxes that deliver excellent performance. In this work, we aim at alleviating this opaqueness of CNNs by providing visual explanations for the network's predictions. Our approach can analyze variety of CNN based models trained for vision applications such as object recognition and caption generation. Unlike existing methods, we achieve this via unraveling the forward pass operation. Proposed method exploits feature dependencies across the layer hierarchy and uncovers the discriminative image locations that guide the network's predictions. We name these locations CNN-Fixations, loosely analogous to human eye fixations. Our approach is a generic method that requires no architectural changes, additional training or gradient computation and computes the important image locations (CNN Fixations). We demonstrate through a variety of applications that our approach is able to localize the discriminative image locations across different network architectures, diverse vision tasks and data modalities.

📄 PDF Abstract BibTeX arXiv:1708.06670

Code (2)

utsavgarg/cnn-fixations 공식 구현 tf
val-iisc/cnn-fixations 공식 구현 tf

Tasks

Caption GenerationImage CaptioningObject Recognition

Similar Papers 제목 키워드 기반

From Semantic Categories to Fixations: A Novel Weakly-Supervised Visual-Auditory Saliency Detection Approach

2021-06-19 · CVPR 2021 1 · Guotao Wang, Chenglizhao Chen, Deng-Ping Fan, Aimin Hao 외

Thanks to the rapid advances in the deep learning techniques and the wide availability of large-scale training sets, the performances of video saliency detection models have been improving steadily and significantly.…

Saliency DetectionVideo Saliency Detection

Look Twice: A Generalist Computational Model Predicts Return Fixations across Tasks and Species

2021-01-05 · Mengmi Zhang, Marcelo Armendariz, Will Xiao, Olivia Rose 외

Primates constantly explore their surroundings via saccadic eye movements that bring different parts of an image into high resolution. In addition to exploring new regions in the visual field, primates also make frequent…

Object Recognition

Understanding Low- and High-Level Contributions to Fixation Prediction

2017-10-01 · ICCV 2017 10 · Matthias Kummerer, Thomas S. A. Wallis, Leon A. Gatys, Matthias Bethge

Understanding where people look in images is an important problem in computer vision. Despite significant research, it remains unclear to what extent human fixations can be predicted by low-level (contrast) compared to h…

Object RecognitionPredictionSaliency PredictionVocal Bursts Intensity Prediction

Visual Decoding of Targets During Visual Search From Human Eye Fixations

2017-06-19 · Hosnieh Sattar, Mario Fritz, Andreas Bulling

What does human gaze reveal about a users' intents and to which extend can these intents be inferred or even visualized? Gaze was proposed as an implicit source of information to predict the target of visual search and, …

Learning of Proto-object Representations via Fixations on Low Resolution

2014-12-23 · Chengyao Shen, Xun Huang, Qi Zhao

While previous researches in eye fixation prediction typically rely on integrating low-level features (e.g. color, edge) to form a saliency map, recently it has been found that the structural organization of these featur…

Object