paper-with-me

Papers

Does Thought Require Sensory Grounding? From Pure Thinkers to Large Language Models

2024-08-18 · David J. Chalmers

Does the capacity to think require the capacity to sense? A lively debate on this topic runs throughout the history of philosophy and now animates discussions of artificial intelligence. I argue that in principle, there can be pure thinkers: thinkers that lack the capacity to sense altogether. I also argue for significant limitations in just what sort of thought is possible in the absence of the capacity to sense. Regarding AI, I do not argue directly that large language models can think or understand, but I rebut one important argument (the argument from sensory grounding) that they cannot. I also use recent results regarding language models to address the question of whether or how sensory grounding enhances cognitive capacities.

📄 PDF Abstract BibTeX arXiv:2408.09605

Code (0)

등록된 구현이 없습니다.

Tasks

Philosophy

Similar Papers 제목 키워드 기반

Seeing the advantage: visually grounding word embeddings to better capture human semantic knowledge

2022-02-21 · CMCL (ACL) 2022 5 · Danny Merkx, Stefan L. Frank, Mirjam Ernestus

Distributional semantic models capture word-level meaning that is useful in many natural language processing tasks and have even been shown to capture cognitive aspects of word meaning. The majority of these models are p…

Grounded language learningImage RetrievalLearning Semantic RepresentationsVisual Grounding+2

EagleVision: A Dual-Stage Framework with BEV-grounding-based Chain-of-Thought for Spatial Intelligence

2025-12-17 · Jiaxu Wan, Xu Wang, Mengwei Xie, Hang Zhang 외 arxiv

Video-based spatial reasoning -- such as estimating distances, judging directions, or understanding layouts from multiple views -- requires selecting informative frames and, when needed, actively seeking additional viewp…

Reinforcement LearningSpatial Reasoning

Interpretable Latent Spaces for Learning from Demonstration

2018-07-17 · Yordan Hristov, Alex Lascarides, Subramanian Ramamoorthy

Effective human-robot interaction, such as in robot learning from human demonstration, requires the learning agent to be able to ground abstract concepts (such as those contained within instructions) in a corresponding h…

DocVAL: Validated Chain-of-Thought Distillation for Grounded Document VQA

2025-11-27 · Pinaki Prasad Guha Neogi, Ahmad Mohammadshirazi, Ser-Nam Lim, Rajiv Ramnath arxiv

Document visual question answering requires models not only to answer questions correctly, but also to precisely localize answers within complex document layouts. While large vision-language models (VLMs) achieve strong …

Visual Question AnsweringSpatial ReasoningText Detection

LaViT: Aligning Latent Visual Thoughts for Multi-modal Reasoning

2026-01-15 · Linquan Wu, Tianxiang Jiang, Yifei Dong, Haoyu Yang 외 arxiv

Current multimodal latent reasoning often relies on external supervision (e.g., auxiliary images), ignoring intrinsic visual attention dynamics. In this work, we identify a critical Perception Gap in distillation: studen…

Visual GroundingText Generation