paper-with-me

홈 › Papers

Sequential Enumeration in Large Language Models

2025-12-04 · Kuinan Hou, Marco Zorzi, Alberto Testolin arxiv

Reliably counting and generating sequences of items remain a significant challenge for neural networks, including Large Language Models (LLMs). Indeed, although this capability is readily handled by rule-based symbolic systems based on serial computation, learning to systematically deploy counting procedures is difficult for neural models, which should acquire these skills through learning. Previous research has demonstrated that recurrent architectures can only approximately track and enumerate sequences of events, and it remains unclear whether modern deep learning systems, including LLMs, can deploy systematic counting procedures over sequences of discrete symbols. This paper aims to fill this gap by investigating the sequential enumeration abilities of five state-of-the-art LLMs, including proprietary, open-source, and reasoning models. We probe LLMs in sequential naming and production tasks involving lists of letters and words, adopting a variety of prompting instructions to explore the role of chain-of-thought in the spontaneous emerging of counting strategies. We also evaluate open-source models with the same architecture but increasing size to see whether the mastering of counting principles follows scaling laws, and we analyze the embedding dynamics during sequential enumeration to investigate the emergent encoding of numerosity. We find that some LLMs are indeed capable of deploying counting procedures when explicitly prompted to do so, but none of them spontaneously engage in counting when simply asked to enumerate the number of items in a sequence. Our results suggest that, despite their impressive emergent abilities, LLMs cannot yet robustly and systematically deploy counting procedures, highlighting a persistent gap between neural and symbolic approaches to compositional generalization.

📄 PDF Abstract BibTeX arXiv:2512.04727

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Asking a Language Model for Diverse Responses

2025-09-22 · Sergey Troshin, Irina Saparina, Antske Fokkens, Vlad Niculae arxiv

Large language models increasingly rely on explicit reasoning chains and can produce multiple plausible responses for a given context. We study the candidate sampler that produces the set of plausible responses contrasti…

A deep real options policy for sequential service region design and timing

2022-12-30 · Srushti Rath, Joseph Y. J. Chow

As various city agencies and mobility operators navigate toward innovative mobility solutions, there is a need for strategic flexibility in well-timed investment decisions in the design and timing of mobility service reg…

Navigate

Sequential Recurrent Neural Networks for Language Modeling

2017-03-23 · Youssef Oualil, Clayton Greenberg, Mittul Singh, Dietrich Klakow

Feedforward Neural Network (FNN)-based language models estimate the probability of the next word based on the history of the last N words, whereas Recurrent Neural Networks (RNN) perform the same task based only on the l…

Language ModelingLanguage ModellingText Compression

History Filtering in Imperfect Information Games: Algorithms and Complexity

2023-11-24 · NeurIPS 2023 11

Historically applied exclusively to perfect information games, depth-limited search with value functions has been key to recent advances in AI for imperfect information games. Most prominent approaches with strong theore…

Card GamesDecision MakingSequential Decision Making

Automated Facility Enumeration for Building Compliance Checking using Door Detection and Large Language Models

2025-09-21 · Licheng Zhang, Bach Le, Naveed Akhtar, Tuan Ngo arxiv

Building compliance checking (BCC) is a critical process for ensuring that constructed facilities meet regulatory standards. A core component of BCC is the accurate enumeration of facility types and their spatial distrib…