paper-with-me

Papers

Evaluating Visual Number Discrimination in Deep Neural Networks

2023-03-13 · Ivana Kajić, Aida Nematzadeh

The ability to discriminate between large and small quantities is a core aspect of basic numerical competence in both humans and animals. In this work, we examine the extent to which the state-of-the-art neural networks designed for vision exhibit this basic ability. Motivated by studies in animal and infant numerical cognition, we use the numerical bisection procedure to test number discrimination in different families of neural architectures. Our results suggest that vision-specific inductive biases are helpful in numerosity discrimination, as models with such biases have lowest test errors on the task, and often have psychometric curves that qualitatively resemble those of humans and animals performing the task. However, even the strongest models, as measured on standard metrics of performance, fail to discriminate quantities in transfer experiments with differing training and testing conditions, indicating that such inductive biases might not be sufficient.

📄 PDF Abstract BibTeX arXiv:2303.07172

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

fail 설명 없음
Test 설명 없음

Similar Papers 제목 키워드 기반

Evaluating Probabilistic Classifiers: The Triptych

2023-01-25 · Timo Dimitriadis, Tilmann Gneiting, Alexander I. Jordan, Peter Vogel

Probability forecasts for binary outcomes, often referred to as probabilistic classifiers or confidence scores, are ubiquitous in science and society, and methods for evaluating and comparing them are in great demand. We…

Diagnostic

Number detectors spontaneously emerge in a deep neural network designed for visual object recognition

2019-05-08 · journal 2019 5 · Khaled Nasr, Pooja Viswanathan, Andreas Nieder

Humans and animals have a “number sense,” an innate capability to intuitively assess the number of visual items in a set, its numerosity. This capability implies that mechanisms to extract numerosity indwell the brain’…

ObjectObject Recognition

The Perceptimatic English Benchmark for Speech Perception Models

2020-05-07 · Juliette Millet, Ewan Dunbar

We present the Perceptimatic English Benchmark, an open experimental benchmark for evaluating quantitative models of speech perception in English. The benchmark consists of ABX stimuli along with the responses of 91 Amer…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Discrimination in the Venture Capital Industry: Evidence from Field Experiments

2020-10-30 · Ye Zhang

This paper examines discrimination by early-stage investors based on startup founders' gender and race using two complementary field experiments with real U.S. venture capitalists. Results show the following. (i) Discrim…

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

2026-04-13 · Xuefeng Wei, Zhixuan Wang, Xuan Zhou, Zhi Qu 외 arxiv

We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and QA. CARTBENCH comprises four subtasks: CURATORQA for evidence-grounde…