paper-with-me

홈 › Papers

Evolution and compression in LLMs: On the emergence of human-aligned categorization

2025-09-09 · Nathaniel Imel, Noga Zaslavsky arxiv

Converging evidence suggests that human systems of semantic categories achieve near-optimal compression via the Information Bottleneck (IB) complexity-accuracy tradeoff. Large language models (LLMs) are not trained for this objective, which raises the question: are LLMs capable of evolving efficient human-aligned semantic systems? To address this question, we focus on color categorization -- a key testbed of cognitive theories of categorization with uniquely rich human data -- and replicate with LLMs two influential human studies. First, we conduct an English color-naming study, showing that LLMs vary widely in their complexity and English-alignment, with larger instruction-tuned models achieving better alignment and IB-efficiency. Second, to test whether these LLMs simply mimic patterns in their training data or actually exhibit a human-like inductive bias toward IB-efficiency, we simulate cultural evolution of pseudo color-naming systems in LLMs via a method we refer to as Iterated in-Context Language Learning (IICLL). We find that akin to humans, LLMs iteratively restructure initially random systems towards greater IB-efficiency. However, only a model with strongest in-context capabilities (Gemini 2.0) is able to recapitulate the wide range of near-optimal IB-tradeoffs observed in humans, while other state-of-the-art models converge to low-complexity solutions. These findings demonstrate how human-aligned semantic categories can emerge in LLMs via the same fundamental principle that underlies semantic efficiency in humans.

📄 PDF Abstract BibTeX arXiv:2509.08093

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Geometric Prior Based Deep Human Point Cloud Geometry Compression

2023-05-02 · Xinju Wu, Pingping Zhang, Meng Wang, Peilin Chen 외

The emergence of digital avatars has raised an exponential increase in the demand for human point clouds with realistic and intricate details. The compression of such data becomes challenging with overwhelming data amoun…

From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning

2025-05-21 · Chen Shani, Dan Jurafsky, Yann Lecun, Ravid Shwartz-Ziv

Humans organize knowledge into compact categories through semantic compression by mapping diverse instances to abstract representations while preserving meaning (e.g., robin and blue jay are both birds; most birds can fl…

Semantic Compression

EFPC: Towards Efficient and Flexible Prompt Compression

2025-03-11 · Yun-Hao Cao, Yangsong Wang, Shuzheng Hao, Zhenxing Li 외

The emergence of large language models (LLMs) like GPT-4 has revolutionized natural language processing (NLP), enabling diverse, complex tasks. However, extensive token counts lead to high computational and financial bur…

S3-CoT: Self-Sampled Succinct Reasoning Enables Efficient Chain-of-Thought LLMs

2026-02-02 · Yanrui Du, Sendong Zhao, Yibo Gao, Danyang Zhao 외 arxiv

Large language models (LLMs) equipped with chain-of-thought (CoT) achieve strong performance and offer a window into LLM behavior. However, recent evidence suggests that improvements in CoT capabilities often come with r…

Domain Generalization

Emergence of more contagious COVID-19 variants from the coevolution of viruses and policy interventions

2021-03-26 · Aymeric Vie

At the end of 2020, policy responses to the SARS-CoV-2 outbreak have been shaken by the emergence of virus variants. The emergence of these more contagious, more severe, or even vaccine-resistant strains have challenged …