paper-with-me

홈 › Papers

Making sense of sensory input

2019-10-05 · Richard Evans, Jose Hernandez-Orallo, Johannes Welbl, Pushmeet Kohli, Marek Sergot

This paper attempts to answer a central question in unsupervised learning: what does it mean to "make sense" of a sensory sequence? In our formalization, making sense involves constructing a symbolic causal theory that both explains the sensory sequence and also satisfies a set of unity conditions. The unity conditions insist that the constituents of the causal theory -- objects, properties, and laws -- must be integrated into a coherent whole. On our account, making sense of sensory input is a type of program synthesis, but it is unsupervised program synthesis. Our second contribution is a computer implementation, the Apperception Engine, that was designed to satisfy the above requirements. Our system is able to produce interpretable human-readable causal theories from very small amounts of data, because of the strong inductive bias provided by the unity conditions. A causal theory produced by our system is able to predict future sensor readings, as well as retrodict earlier readings, and impute (fill in the blanks of) missing sensory readings, in any combination. We tested the engine in a diverse variety of domains, including cellular automata, rhythms and simple nursery tunes, multi-modal binding problems, occlusion tasks, and sequence induction intelligence tests. In each domain, we test our engine's ability to predict future sensor values, retrodict earlier sensor values, and impute missing sensory data. The engine performs well in all these domains, significantly out-performing neural net baselines. We note in particular that in the sequence induction intelligence tests, our system achieved human-level performance. This is notable because our system is not a bespoke system designed specifically to solve intelligence tests, but a general-purpose system that was designed to make sense of any sensory sequence.

📄 PDF Abstract BibTeX arXiv:1910.02227

Code (1)

RichardEvans/apperception 공식 구현

Tasks

Inductive BiasProgram SynthesisUnity

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Task-Oriented Mulsemedia Communication using Unified Perceiver and Conformal Prediction in 6G Wireless Systems

2024-05-14 · Hongzhi Guo, Ian F. Akyildiz

The growing prominence of eXtended Reality (XR), holographic-type communications, and metaverse demands truly immersive user experiences by using many sensory modalities, including sight, hearing, touch, smell, taste, et…

Conformal Prediction

Sense it with your eyes: Sensation Generation and Understanding for Advertisements

2026-07-28 · Aysan Aghazadeh, Sina Malakouti, Adriana Kovashka arxiv

Sensory advertising evokes human senses through visual cues, enabling audiences to mentally simulate experiences and increasing persuasive impact. Despite the recent increase in using AI in generating and understanding c…

FoodSense: A Multisensory Food Dataset and Benchmark for Predicting Taste, Smell, Texture, and Sound from Images

2026-04-15 · Sabab Ishraq, Aarushi Aarushi, Juncai Jiang, Chen Chen arxiv

Humans routinely infer taste, smell, texture, and even sound from food images a phenomenon well studied in cognitive science. However, prior vision language research on food has focused primarily on recognition tasks suc…

Sensory-driven microinterventions for improved health and wellbeing

2025-03-18 · Youssef Abdalla, Elia Gatti, Mine Orlu, Marianna Obrist

The five senses are gateways to our wellbeing and their decline is considered a significant public health challenge which is linked to multiple conditions that contribute significantly to morbidity and mortality. Modern …

Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks

2018-10-24 · Michelle A. Lee, Yuke Zhu, Krishnan Srinivasan, Parth Shah 외

Contact-rich manipulation tasks in unstructured environments often require both haptic and visual feedback. However, it is non-trivial to manually design a robot controller that combines modalities with very different ch…

Contact-rich ManipulationDeep Reinforcement LearningReinforcement LearningReinforcement Learning (RL)+1