paper-with-me

홈 › Papers

Guidance is All You Need: Temperature-Guided Reasoning in Large Language Models

2024-12-05 · Eyad Gomaa, Gomaa Salah

We present Quasar-1, a novel architecture that introduces temperature-guided reasoning to large language models through the Token Temperature Mechanism (TTM) and Guided Sequence of Thought (GSoT). Our approach leverages the concept of hot and cold tokens, where hot tokens are prioritized for their contextual relevance, while cold tokens provide supplementary information. This dynamic modulation of token importance enables the model to achieve superior logical reasoning capabilities compared to traditional chain-of-thought approaches. Through rigorous mathematical analysis, we prove that our temperature-guided attention mechanism converges to optimal reasoning paths with exponential guarantees. Empirical results show significant improvements in reasoning accuracy and computational efficiency across a wide range of tasks, making advanced AI reasoning accessible to a broader range of applications.

📄 PDF Abstract BibTeX arXiv:2412.06822

Code (0)

등록된 구현이 없습니다.

Tasks

AllComputational EfficiencyLogical Reasoning

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Self-Evaluation Guided Beam Search for Reasoning

2023-05-01 · NeurIPS 2023 11 · Yuxi Xie, Kenji Kawaguchi, Yiran Zhao, Xu Zhao 외

Breaking down a problem into intermediate steps has demonstrated impressive performance in Large Language Model (LLM) reasoning. However, the growth of the reasoning chain introduces uncertainty and error accumulation, m…

Arithmetic ReasoningGSM8KLanguage ModelingLanguage Modelling+2

ReconMOST: Multi-Layer Sea Temperature Reconstruction with Observations-Guided Diffusion

2025-06-12 · Yuanyi Song, Pumeng Lyu, Ben Fei, Fenghua Ling 외

Accurate reconstruction of ocean is essential for reflecting global climate dynamics and supporting marine meteorological research. Conventional methods face challenges due to sparse data, algorithmic complexity, and hig…

G$^2$RPO-A: Guided Group Relative Policy Optimization with Adaptive Guidance

2025-08-18 · Yongxin Guo, Wenbo Deng, Zhenglin Cheng, Xiaoying Tang arxiv

Reinforcement Learning with Verifiable Rewards (RLVR) has markedly enhanced the reasoning abilities of large language models (LLMs). Its success, however, largely depends on strong base models with rich world knowledge, …

Reinforcement LearningMathematical Reasoning

MRI-Guided High Intensity Focused Ultrasound of Liver and Kidney

2020-11-21 · Baudouin Denis de Senneville, Mario Ries, Wilbert Bartels, Chrit Moonen

High Intensity Focused Ultrasound (HIFU) can be used to achieve a local temperature increase deep inside the human body in a non-invasive way. MRI guidance of the procedure allows in situ target definition. In addition, …

Motion CompensationVocal Bursts Intensity Prediction

Structured Chemistry Reasoning with Large Language Models

2023-11-16 · Siru Ouyang, Zhuosheng Zhang, Bing Yan, Xuan Liu 외

Large Language Models (LLMs) excel in diverse areas, yet struggle with complex scientific reasoning, especially in the field of chemistry. Different from the simple chemistry tasks (e.g., molecule classification) address…

General Knowledge