paper-with-me

홈 › Papers

Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games

2026-03-23 · Clément Hongler, Franck Gabriel, Valentin Hartmann, Arthur Renard, Andrew Emil arxiv

Defining a constructive process to build general capabilities for language models in an automatic manner is considered an open problem in artificial intelligence. Towards this, we consider the problem of building a curriculum of tasks that grows a model via relevant skill discovery. We provide a concrete framework for this task, using a family of tasks called Cross-Entropy Games, which we postulate is universal in a suitable sense. We show that if it is possible to grow the curriculum for relevant skill discovery by iterating a greedy optimization algorithm, then, under natural assumptions, there is essentially only one meta-objective possible (up to a few hyper-parameters). We call the resulting process cognitive training. We postulate that, given sufficiently capable language models as players and meta-samplers, cognitive training provides a principled way to relevant skill discovery; and hence to the extent general capabilities are achievable via greedy curriculum learning, cognitive training would be a solution.

📄 PDF Abstract BibTeX arXiv:2603.22479

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving

2025-06-10 · Yuxuan Zhou, Xien Liu, Chenwei Yan, Chen Ning 외

Large language models (LLMs) have demonstrated remarkable performance on various medical benchmarks, but their capabilities across different cognitive levels remain underexplored. Inspired by Bloom's Taxonomy, we propose…

Scaling Auditory Cognition via Test-Time Compute in Audio Language Models

2025-03-30 · Ting Dang, Yan Gao, Hong Jia

Large language models (LLMs) have shown exceptional versatility in natural language processing, prompting recent efforts to extend their multimodal capabilities to speech processing through the development of audio large…

speech-recognitionSpeech Recognition

Human-like Cognitive Generalization for Large Models via Brain-in-the-loop Supervision

2025-05-14 · Jiaxuan Chen, Yu Qi, Yueming Wang, Gang Pan

Recent advancements in deep neural networks (DNNs), particularly large-scale language models, have demonstrated remarkable capabilities in image and natural language understanding. Although scaling up model parameters wi…

Natural Language UnderstandingZero-Shot Learning

Line Goes Up? Inherent Limitations of Benchmarks for Evaluating Large Language Models

2025-02-20 · James Fodor

Large language models (LLMs) regularly demonstrate new and impressive performance on a wide range of language, knowledge, and reasoning benchmarks. Such rapid progress has led many commentators to argue that LLM general …

Benchmarking

Define, Evaluate, and Improve Task-Oriented Cognitive Capabilities for Instruction Generation Models

2022-12-21 · Lingjun Zhao, Khanh Nguyen, Hal Daumé III

Recent work studies the cognitive capabilities of language models through psychological tests designed for humans. While these studies are helpful for understanding the general capabilities of these models, there is no g…

Language Modelling