paper-with-me

홈 › Papers

Capturing the objects of vision with neural networks

2021-09-07 · Benjamin Peters, Nikolaus Kriegeskorte

Human visual perception carves a scene at its physical joints, decomposing the world into objects, which are selectively attended, tracked, and predicted as we engage our surroundings. Object representations emancipate perception from the sensory input, enabling us to keep in mind that which is out of sight and to use perceptual content as a basis for action and symbolic cognition. Human behavioral studies have documented how object representations emerge through grouping, amodal completion, proto-objects, and object files. Deep neural network (DNN) models of visual object recognition, by contrast, remain largely tethered to the sensory input, despite achieving human-level performance at labeling objects. Here, we review related work in both fields and examine how these fields can help each other. The cognitive literature provides a starting point for the development of new experimental tasks that reveal mechanisms of human object perception and serve as benchmarks driving development of deep neural network models that will put the object into object recognition.

📄 PDF Abstract BibTeX arXiv:2109.03351

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectObject Recognition

Similar Papers 제목 키워드 기반

Capturing and Recognizing Objects Appearance Employing Eigenspace

2014-03-25 · M. Ashrafuzzaman, M. M . Rahman, M. M. A. Hashem

This paper presents a method of capturing objects appearances from its environment and it also describes how to recognize unknown appearances creating an eigenspace. This representation and recognition can be done automa…

Pose-Invariant Object Recognition for Event-Based Vision with Slow-ELM

2019-03-19 · Rohan Ghosh, Siyi Tang, Mahdi Rasouli, Nitish Thakor 외

Neuromorphic image sensors produce activity-driven spiking output at every pixel. These low-power consuming imagers which encode visual change information in the form of spikes help reduce computational overhead and real…

CPUEvent-based visionGeneral ClassificationObject Recognition+1

Context-Aware Semantic Segmentation: Enhancing Pixel-Level Understanding with Large Language Models for Advanced Vision Applications

2025-03-25 · Ben Rahman

Semantic segmentation has made significant strides in pixel-level image understanding, yet it remains limited in capturing contextual and semantic relationships between objects. Current models, such as CNN and Transforme…

Autonomous DrivingSemantic Segmentation

DexterCap: An Affordable and Automated System for Capturing Dexterous Hand-Object Manipulation

2026-01-09 · Yutong Liang, Shiyi Xu, Yulong Zhang, Bowen Zhan 외 arxiv

Capturing fine-grained hand-object interactions is challenging due to severe self-occlusion from closely spaced fingers and the subtlety of in-hand manipulation motions. Existing optical motion capture systems rely on ex…

Refractive Light-Field Features for Curved Transparent Objects in Structure from Motion

2021-03-29 · Dorian Tsai, Peter Corke, Thierry Peynot, Donald G. Dansereau

Curved refractive objects are common in the human environment, and have a complex visual appearance that can cause robotic vision algorithms to fail. Light-field cameras allow us to address this challenge by capturing th…

Transparent objects