paper-with-me

홈 › Papers

Beyond the Black Box: A Survey on the Theory and Mechanism of Large Language Models

2026-01-06 · Zeyu Gan, Ruifeng Ren, Wei Yao, Xiaolin Hu, Gengze Xu, Chen Qian, Huayi Tang, Zixuan Gong, Xinhao Yao, Pengwei Tang, Zhenxing Dou, Yong Liu arxiv

The rapid emergence of Large Language Models (LLMs) has precipitated a profound paradigm shift in Artificial Intelligence, delivering monumental engineering successes that increasingly impact modern society. However, a critical paradox persists within the current field: despite the empirical efficacy, our theoretical understanding of LLMs remains disproportionately nascent, forcing these systems to be treated largely as ``black boxes''. To address this theoretical fragmentation, this survey proposes a unified lifecycle-based taxonomy that organizes the research landscape into six distinct stages: Data Preparation, Model Preparation, Training, Alignment, Inference, and Evaluation. Within this framework, we provide a systematic review of the foundational theories and internal mechanisms driving LLM performance. Specifically, we analyze core theoretical issues such as the mathematical justification for data mixtures, the representational limits of various architectures, and the optimization dynamics of alignment algorithms. Moving beyond current best practices, we identify critical frontier challenges, including the theoretical limits of synthetic data self-improvement, the mathematical bounds of safety guarantees, and the mechanistic origins of emergent intelligence. By connecting empirical observations with rigorous scientific inquiry, this work provides a structured roadmap for transitioning LLM development from engineering heuristics toward a principled scientific discipline.

📄 PDF Abstract BibTeX arXiv:2601.02907

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Theoretical Survey on Foundation Models

2024-10-15 · Shi Fu, Yuzhu Chen, Yingjie Wang, DaCheng Tao

Understanding the inner mechanisms of black-box foundation models (FMs) is essential yet challenging in artificial intelligence and its applications. Over the last decade, the long-running focus has been on their explain…

Learning TheorySurvey

Opening the Black Box: A Survey on the Mechanisms of Multi-Step Reasoning in Large Language Models

2026-01-02 · Liangming Pan, Jason Liang, Jiaran Ye, Minglai Yang 외 arxiv

Large Language Models (LLMs) have demonstrated remarkable abilities to solve problems requiring multiple reasoning steps, yet the internal mechanisms enabling such capabilities remain elusive. Unlike existing surveys tha…

The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence

2026-04-08 · Napoleon Paxton arxiv

Dr. David Blackwell was a mathematician and statistician of the first rank, whose contributions to statistical theory, game theory, and decision theory predated many of the algorithmic breakthroughs that define modern ar…

Reinforcement LearningRobot NavigationDecision Making

Complexity Theory for Discrete Black-Box Optimization Heuristics

2018-01-06 · Carola Doerr

A predominant topic in the theory of evolutionary algorithms and, more generally, theory of randomized black-box optimization techniques is running time analysis. Running time analysis aims at understanding the performan…

Evolutionary Algorithms

XChoice: Explainable Evaluation of AI-Human Alignment in LLM-based Constrained Choice Decision Making

2026-01-16 · Weihong Qi, Fan Huang, Rasika Muralidharan, Jisun An 외 arxiv

We present XChoice, an explainable framework for evaluating AI-human alignment in constrained decision making. Moving beyond outcome agreement such as accuracy and F1 score, XChoice fits a mechanism-based decision model …

Decision Making