paper-with-me

홈 › Papers

A model for full local image interpretation

2021-10-17 · Guy Ben-Yosef, Liav Assif, Daniel Harari, Shimon Ullman

We describe a computational model of humans' ability to provide a detailed interpretation of components in a scene. Humans can identify in an image meaningful components almost everywhere, and identifying these components is an essential part of the visual process, and of understanding the surrounding scene and its potential meaning to the viewer. Detailed interpretation is beyond the scope of current models of visual recognition. Our model suggests that this is a fundamental limitation, related to the fact that existing models rely on feed-forward but limited top-down processing. In our model, a first recognition stage leads to the initial activation of class candidates, which is incomplete and with limited accuracy. This stage then triggers the application of class-specific interpretation and validation processes, which recover richer and more accurate interpretation of the visible scene. We discuss implications of the model for visual interpretation by humans and by computer vision models.

📄 PDF Abstract BibTeX arXiv:2110.08744

Code (0)

등록된 구현이 없습니다.

Tasks

model

Similar Papers 제목 키워드 기반

Full interpretation of minimal images

2018-02-01 · Guy Ben-Yosef; Liav Assif; Shimon Ullman

The goal in this work is to model the process of ‘full interpretation’ of object images, which is the ability to identify and localize all semantic features and parts that are recognized by human observers. The task is a…

Structured learning and detailed interpretation of minimal object images

2017-11-29 · Guy Ben-Yosef, Liav Assif, Shimon Ullman

We model the process of human full interpretation of object images, namely the ability to identify and localize all semantic features and parts that are recognized by human observers. The task is approached by dividing t…

SkySense-O: Towards Open-World Remote Sensing Interpretation with Vision-Centric Visual-Language Modeling

2025-01-01 · CVPR 2025 1 · Qi Zhu, Jiangwei Lao, Deyi Ji, Junwei Luo 외

Open-world interpretation aims to accurately localize and recognize all objects within images by vision-language models (VLMs). While substantial progress has been made in this task for natural images, the advancemen…

Language ModelingLanguage Modelling

Can I trust you more? Model-Agnostic Hierarchical Explanations

2018-12-12 · ICLR 2019 5 · Michael Tsang, Youbang Sun, Dongxu Ren, Yan Liu

Interactions such as double negation in sentences and scene interactions in images are common forms of complex dependencies captured by state-of-the-art machine learning models. We propose Mah\'e, a novel approach to pro…

BIG-bench Machine LearningNegation

Understanding Interpretation Difficulty in Harmful Online Communication: Insights from Cybercrime Communities

2026-07-08 · Tomohiro Okatsu, Naoki Takada, Yin Min Pa Pa, Katsunari Yoshioka 외 arxiv

Harmful online communication often contains slang, coded terms, abbreviations, and community-specific expressions, which make messages difficult to interpret. This paper presents an exploratory study of interpretation di…