paper-with-me

Papers

Human-like object concept representations emerge naturally in multimodal large language models

2024-07-01 · Changde Du, Kaicheng Fu, Bincheng Wen, Yi Sun, Jie Peng, Wei Wei, Ying Gao, Shengpei Wang, Chuncheng Zhang, Jinpeng Li, Shuang Qiu, Le Chang, Huiguang He

Understanding how humans conceptualize and categorize natural objects offers critical insights into perception and cognition. With the advent of Large Language Models (LLMs), a key question arises: can these models develop human-like object representations from linguistic and multimodal data? In this study, we combined behavioral and neuroimaging analyses to explore the relationship between object concept representations in LLMs and human cognition. We collected 4.7 million triplet judgments from LLMs and Multimodal LLMs (MLLMs) to derive low-dimensional embeddings that capture the similarity structure of 1,854 natural objects. The resulting 66-dimensional embeddings were stable, predictive, and exhibited semantic clustering similar to human mental representations. Remarkably, the dimensions underlying these embeddings were interpretable, suggesting that LLMs and MLLMs develop human-like conceptual representations of objects. Further analysis showed strong alignment between model embeddings and neural activity patterns in brain regions such as EBA, PPA, RSC, and FFA. This provides compelling evidence that the object representations in LLMs, while not identical to human ones, share fundamental similarities that reflect key aspects of human conceptual knowledge. Our findings advance the understanding of machine intelligence and inform the development of more human-like artificial cognitive systems.

📄 PDF Abstract BibTeX arXiv:2407.01067

Code (0)

등록된 구현이 없습니다.

Tasks

Triplet

Similar Papers 제목 키워드 기반

Human-like conceptual representations emerge from language prediction

2025-01-21 · Ningyu Xu, Qi Zhang, Chao Du, Qiang Luo 외

People acquire concepts through rich physical and social experiences and use them to understand the world. In contrast, large language models (LLMs), trained exclusively through next-token prediction over language data, …

PredictionReverse Dictionary

How can embedding models bind concepts?

2026-05-29 · Arnas Uselis, Darina Koishigarina, Seong Joon Oh arxiv

Humans easily determine which color belongs to which shape in multi-object scenes, an ability known as concept binding. Vision-language embedding models such as CLIP struggle with binding: they recognize individual conce…

Cross-Modal Retrieval

A neural network for modeling human concept formation, understanding and communication

2026-01-05 · Liangxuan Guo, Haoyang Chen, Yang Chen, Yanchao Bi 외 arxiv

A remarkable capability of the human brain is to form more abstract conceptual representations from sensorimotor experiences and flexibly apply them independent of direct sensory inputs. However, the computational mechan…

Emergent Communication with Attention

2023-05-18 · Ryokan Ri, Ryo Ueda, Jason Naradowsky

To develop computational agents that better communicate using their own emergent language, we endow the agents with an ability to focus their attention on particular concepts in the environment. Humans often understand a…

Few-Shot Learning of Visual Compositional Concepts through Probabilistic Schema Induction

2025-05-14 · Andrew Jun Lee, Taylor Webb, Trevor Bihl, Keith Holyoak 외

The ability to learn new visual concepts from limited examples is a hallmark of human cognition. While traditional category learning models represent each example as an unstructured feature vector, compositional concept …

Deep LearningFew-Shot Learning