paper-with-me

홈 › Papers

Putting visual object recognition in context

2019-11-17 · CVPR 2020 6 · Mengmi Zhang, Claire Tseng, Gabriel Kreiman

Context plays an important role in visual recognition. Recent studies have shown that visual recognition networks can be fooled by placing objects in inconsistent contexts (e.g., a cow in the ocean). To model the role of contextual information in visual recognition, we systematically investigated ten critical properties of where, when, and how context modulates recognition, including the amount of context, context and object resolution, geometrical structure of context, context congruence, and temporal dynamics of contextual modulation. The tasks involved recognizing a target object surrounded with context in a natural image. As an essential benchmark, we conducted a series of psychophysics experiments where we altered one aspect of context at a time, and quantified recognition accuracy. We propose a biologically-inspired context-aware object recognition model consisting of a two-stream architecture. The model processes visual information at the fovea and periphery in parallel, dynamically incorporates object and contextual information, and sequentially reasons about the class label for the target object. Across a wide range of behavioral tasks, the model approximates human level performance without retraining for each task, captures the dependence of context enhancement on image properties, and provides initial steps towards integrating scene and object information for visual recognition. All source code and data are publicly available: https://github.com/kreimanlab/Put-In-Context.

📄 PDF Abstract BibTeX arXiv:1911.07349

Code (1)

kreimanlab/Put-In-Context 공식 구현 pytorch

Tasks

ObjectObject Recognition

Similar Papers 제목 키워드 기반

Fine-grained Activities of People Worldwide

2022-07-11 · Jeffrey Byrne, Greg Castanon, Zhongheng Li, Gil Ettinger

Every day, humans perform many closely related activities that involve subtle discriminative motions, such as putting on a shirt vs. putting on a jacket, or shaking hands vs. giving a high five. Activity recognition by e…

Action DetectionActivity DetectionActivity RecognitionDiversity

Convolutional Neural Networks with Gated Recurrent Connections

2021-06-05 · JianFeng Wang, Xiaolin Hu

The convolutional neural network (CNN) has become a basic model for solving many computer vision problems. In recent years, a new class of CNNs, recurrent convolution neural network (RCNN), inspired by abundant recurrent…

object-detectionObject DetectionObject RecognitionScene Text Recognition

Evaluation of Large Language Models for Decision Making in Autonomous Driving

2023-12-11 · Kotaro Tanahashi, Yuichi Inoue, Yu Yamaguchi, Hidetatsu Yaginuma 외

Various methods have been proposed for utilizing Large Language Models (LLMs) in autonomous driving. One strategy of using LLMs for autonomous driving involves inputting surrounding objects as text prompts to the LLMs, a…

Autonomous DrivingDecision Making

Exploring Context and Visual Pattern of Relationship for Scene Graph Generation

2019-06-01 · CVPR 2019 6 · Wenbin Wang, Ruiping Wang, Shiguang Shan, Xilin Chen

Relationship is the core of scene graph, but its prediction is far from satisfying because of its complex visual diversity. To alleviate this problem, we treat relationship as an abstract object, exploring not only signi…

DiversityGraph GenerationObjectObject Recognition+1

The Effect of Top-Down Attention in Occluded Object Recognition

2020-07-17 · Zahra Sadeghi

This study is concerned with the top-down visual processing benefit in the task of occluded object recognition. To this end, a psychophysical experiment is designed and carried out which aimed at investigating the effect…

ObjectObject Recognition