paper-with-me

홈 › Papers

Lines of Thought in Large Language Models

2024-10-02 · Raphaël Sarfati, Toni J. B. Liu, Nicolas Boullé, Christopher J. Earls

Large Language Models achieve next-token prediction by transporting a vectorized piece of text (prompt) across an accompanying embedding space under the action of successive transformer layers. The resulting high-dimensional trajectories realize different contextualization, or 'thinking', steps, and fully determine the output probability distribution. We aim to characterize the statistical properties of ensembles of these 'lines of thought.' We observe that independent trajectories cluster along a low-dimensional, non-Euclidean manifold, and that their path can be well approximated by a stochastic equation with few parameters extracted from data. We find it remarkable that the vast complexity of such large models can be reduced to a much simpler form, and we reflect on implications.

📄 PDF Abstract BibTeX arXiv:2410.01545

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Latent Thought Flow: Efficient Latent Reasoning in Large Language Models

2026-06-15 · Xiandong Zou, Jing Huang, Jianshu Li, Pan Zhou arxiv

Large Language Models (LLMs) increasingly rely on intermediate reasoning, yet explicit Chain-of-Thought (CoT) suffers from a linguistic space bottleneck: each thought must be decoded into tokens, causing high inference o…

Transfer Learning

Table as Thought: Exploring Structured Thoughts in LLM Reasoning

2025-01-04 · Zhenjie Sun, Naihao Deng, Haofei Yu, Jiaxuan You

Large language models' reasoning abilities benefit from methods that organize their thought processes, such as chain-of-thought prompting, which employs a sequential structure to guide the reasoning process step-by-step.…

Mathematical Reasoning

Prompting Large Language Models with Rationale Heuristics for Knowledge-based Visual Question Answering

2024-12-22 · Zhongjian Hu, Peng Yang, Bing Li, Fengyuan Liu

Recently, Large Language Models (LLMs) have been used for knowledge-based Visual Question Answering (VQA). Despite the encouraging results of previous studies, prior methods prompt LLMs to predict answers directly, negle…

Question AnsweringVisual Question AnsweringVisual Question Answering (VQA)

MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning

2024-09-18 · Ningyuan Xi, Xiaoyu Wang, Yetao Wu, Teng Chen 외

Large Language Model can reasonably understand and generate human expressions but may lack of thorough thinking and reasoning mechanisms. Recently there have been several studies which enhance the thinking ability of lan…

Language ModelingLanguage ModellingLarge Language Model

Debiasing Large Language Models via Adaptive Causal Prompting with Sketch-of-Thought

2026-01-13 · Bowen Li, Ziqi Xu, Jing Ren, Renqiang Luo 외 arxiv

Despite notable advancements in prompting methods for Large Language Models (LLMs), such as Chain-of-Thought (CoT), existing strategies still suffer from excessive token usage and limited generalisability across diverse …

Computational Efficiency