paper-with-me

홈 › Papers

Do Language Models Have Bayesian Brains? Distinguishing Stochastic and Deterministic Decision Patterns within Large Language Models

2025-06-12 · Andrea Yaoyun Cui, Pengfei Yu

Language models are essentially probability distributions over token sequences. Auto-regressive models generate sentences by iteratively computing and sampling from the distribution of the next token. This iterative sampling introduces stochasticity, leading to the assumption that language models make probabilistic decisions, similar to sampling from unknown distributions. Building on this assumption, prior research has used simulated Gibbs sampling, inspired by experiments designed to elicit human priors, to infer the priors of language models. In this paper, we revisit a critical question: Do language models possess Bayesian brains? Our findings show that under certain conditions, language models can exhibit near-deterministic decision-making, such as producing maximum likelihood estimations, even with a non-zero sampling temperature. This challenges the sampling assumption and undermines previous methods for eliciting human-like priors. Furthermore, we demonstrate that without proper scrutiny, a system with deterministic behavior undergoing simulated Gibbs sampling can converge to a "false prior." To address this, we propose a straightforward approach to distinguish between stochastic and deterministic decision patterns in Gibbs sampling, helping to prevent the inference of misleading language model priors. We experiment on a variety of large language models to identify their decision patterns under various circumstances. Our results provide key insights in understanding decision making of large language models.

📄 PDF Abstract BibTeX arXiv:2506.10268

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Bayesian Approaches for Revealing Complex Neural Network Dynamics in Parkinson's Disease

2024-10-30 · Hina Shaheen, Roderick Melnik

Parkinson's disease (PD) belongs to the class of neurodegenerative disorders that affect the central nervous system. It is usually defined as the gradual loss of dopaminergic neurons in the substantia nigra pars compacta…

Bayesian Inference

Stochastic Normalizations as Bayesian Learning

2018-11-01 · Alexander Shekhovtsov, Boris Flach

In this work we investigate the reasons why Batch Normalization (BN) improves the generalization performance of deep networks. We argue that one major reason, distinguishing it from data-independent normalization methods…

BrainStorm @ iREL at #SMM4H 2024: Leveraging Translation and Topical Embeddings for Annotation Detection in Tweets

2024-05-18 · Manav Chaudhary, Harshit Gupta, Vasudeva Varma

The proliferation of LLMs in various NLP tasks has sparked debates regarding their reliability, particularly in annotation tasks where biases and hallucinations may arise. In this shared task, we address the challenge of…

The Origin of Inference: Ediacaran Ecology and the Evolution of Bayesian Brains

2015-04-12

The evolution of spiking neurons and nervous systems in the late Ediacaran period simultaneously with the evolution of carnivores around 550 million years ago can be explained by the need for accurately timed decisions u…

Decision Making

Brainstorming Brings Power to Large Language Models of Knowledge Reasoning

2024-06-02 · Zining Qin, Chenhao Wang, Huiling Qin, Weijia Jia

Large Language Models (LLMs) have demonstrated amazing capabilities in language generation, text comprehension, and knowledge reasoning. While a single powerful model can already handle multiple tasks, relying on a singl…

Logical ReasoningReading ComprehensionText Generation