paper-with-me

Papers

How does the primate brain combine generative and discriminative computations in vision?

2024-01-11 · Benjamin Peters, James J. DiCarlo, Todd Gureckis, Ralf Haefner, Leyla Isik, Joshua Tenenbaum, Talia Konkle, Thomas Naselaris, Kimberly Stachenfeld, Zenna Tavares, Doris Tsao, Ilker Yildirim, Nikolaus Kriegeskorte

Vision is widely understood as an inference problem. However, two contrasting conceptions of the inference process have each been influential in research on biological vision as well as the engineering of machine vision. The first emphasizes bottom-up signal flow, describing vision as a largely feedforward, discriminative inference process that filters and transforms the visual information to remove irrelevant variation and represent behaviorally relevant information in a format suitable for downstream functions of cognition and behavioral control. In this conception, vision is driven by the sensory data, and perception is direct because the processing proceeds from the data to the latent variables of interest. The notion of "inference" in this conception is that of the engineering literature on neural networks, where feedforward convolutional neural networks processing images are said to perform inference. The alternative conception is that of vision as an inference process in Helmholtz's sense, where the sensory evidence is evaluated in the context of a generative model of the causal processes giving rise to it. In this conception, vision inverts a generative model through an interrogation of the evidence in a process often thought to involve top-down predictions of sensory data to evaluate the likelihood of alternative hypotheses. The authors include scientists rooted in roughly equal numbers in each of the conceptions and motivated to overcome what might be a false dichotomy between them and engage the other perspective in the realm of theory and experiment. The primate brain employs an unknown algorithm that may combine the advantages of both conceptions. We explain and clarify the terminology, review the key empirical evidence, and propose an empirical research program that transcends the dichotomy and sets the stage for revealing the mysterious hybrid algorithm of primate vision.

📄 PDF Abstract BibTeX arXiv:2401.06005

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unveiling Secrets of Brain Function With Generative Modeling: Motion Perception in Primates & Cortical Network Organization in Mice

2024-12-25 · Hadi Vafaii

This Dissertation is comprised of two main projects, addressing questions in neuroscience through applications of generative modeling. Project #1 (Chapter 4) explores how neurons encode features of the external world. I …

Simple Models, Rich Representations: Visual Decoding from Primate Intracortical Neural Signals

2026-01-16 · Matteo Ciferri, Matteo Ferrante, Nicola Toschi arxiv

Understanding how neural activity gives rise to perception is a central challenge in neuroscience. We address the problem of decoding visual information from high-density intracortical recordings in primates, using the T…

Image Retrieval

Towards Transcranial 3D Ultrasound Localization Microscopy of the Nonhuman Primate Brain

2024-04-04 · Paul Xing, Vincent Perrot, Adan Ulises Dominguez-Vargas, Stephan Quessy 외

Hemodynamic changes occur in stroke and neurodegenerative diseases. Developing imaging techniques allowing the in vivo visualization and quantification of cerebral blood flow would help better understand the underlying m…

Hierarchical VAEs provide a normative account of motion processing in the primate brain

2023-09-21 · NeurIPS 2023 11

The relationship between perception and inference, as postulated by Helmholtz in the 19th century, is paralleled in modern machine learning by generative models like Variational Autoencoders (VAEs) and their hierarchical…

Masked Image Modeling as a Framework for Self-Supervised Learning across Eye Movements

2024-04-12 · Robin Weiler, Matthias Brucklacher, Cyriel M. A. Pennartz, Sander M. Bohté

To make sense of their surroundings, intelligent systems must transform complex sensory inputs to structured codes that are reduced to task-relevant information such as object category. Biological agents achieve this in …

Data AugmentationRepresentation LearningSelf-Supervised Learning