paper-with-me

Papers

Optimality and limitations of audio-visual integration for cognitive systems

2019-12-02 · W. Paul Boyce, Tony Lindsay, Arkady Zgonnikov, Ignacio Rano, KongFatt Wong-Lin

Multimodal integration is an important process in perceptual decision-making. In humans, this process has often been shown to be statistically optimal, or near optimal: sensory information is combined in a fashion that minimises the average error in perceptual representation of stimuli. However, sometimes there are costs that come with the optimization, manifesting as illusory percepts. We review audio-visual facilitations and illusions that are products of multisensory integration, and the computational models that account for these phenomena. In particular, the same optimal computational model can lead to illusory percepts, and we suggest that more studies should be needed to detect and mitigate these illusions, as artefacts in artificial cognitive systems. We provide cautionary considerations when designing artificial cognitive systems with the view of avoiding such artefacts. Finally, we suggest avenues of research towards solutions to potential pitfalls in system design. We conclude that detailed understanding of multisensory integration and the mechanisms behind audio-visual illusions can benefit the design of artificial cognitive systems.

📄 PDF Abstract BibTeX arXiv:1912.00581

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Brain Connectivity Features-based Age Group Classification using Temporal Asynchrony Audio-Visual Integration Task

2023-04-13 · Prerna Singh, Ayush Tripathi, Lalan Kumar, Tapan Kumar Gandhi

The process of integration of inputs from several sensory modalities in the human brain is referred to as multisensory integration. Age-related cognitive decline leads to a loss in the ability of the brain to conceive mu…

EEGFunctional Connectivity

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs

2026-06-01 · Yaoting Wang, Ziyi Zhang, Wenming Tu, Shaoxuan Xu 외 arxiv

Recent advances in Omni-Multimodal Large Language Models (Omni-MLLMs) have enabled strong integration of vision, audio, and language. However, their audio-visual intelligence (AVI) remains insufficiently evaluated due to…

Audiovisual angle and voice incongruence do not affect audiovisual verbal short-term memory in virtual reality

2024-10-30 · Cosima A. Ermert, Manuj Yadav, Jonathan Ehret, Chinthusa Mohanathasan 외

Virtual reality (VR) environments are frequently used in auditory and cognitive research to imitate real-life scenarios, presumably enhancing state-of-the-art approaches with traditional computer screens. However, the ef…

Insights into Age-Related Functional Brain Changes during Audiovisual Integration Tasks: A Comprehensive EEG Source-Based Analysis

2023-11-27 · Prerna Singh, Ayush Tripathi, Lalan Kumar, Tapan Kumar Gandhi

The seamless integration of visual and auditory information is a fundamental aspect of human cognition. Although age-related functional changes in Audio-Visual Integration (AVI) have been extensively explored in the past…

EEGFunctional Connectivity

SAVEn-Vid: Synergistic Audio-Visual Integration for Enhanced Understanding in Long Video Context

2024-11-25 · Jungang Li, Sicheng Tao, Yibo Yan, Xiaojie Gu 외

Endeavors have been made to explore Large Language Models for video analysis (Video-LLMs), particularly in understanding and interpreting long videos. However, existing Video-LLMs still face challenges in effectively int…

Large Language ModelMMEVideo MMEVideo Understanding