paper-with-me

홈 › Papers

ROI-Reasoning: Rational Optimization for Inference via Pre-Computation Meta-Cognition

2026-01-07 · Muyang Zhao, Qi Qi, Hao Sun arxiv

Large language models (LLMs) can achieve strong reasoning performance with sufficient computation, but they do not inherently know how much computation a task requires. We study budgeted inference-time reasoning for multiple tasks under a strict global token constraint and formalize it as a Ordered Stochastic Multiple-Choice Knapsack Problem(OS-MCKP). This perspective highlights a meta-cognitive requirement -- anticipating task difficulty, estimating return over investment (ROI), and allocating computation strategically. We propose ROI-Reasoning, a two-stage framework that endows LLMs with intrinsic, budget-aware rationality. In the first stage, Meta-Cognitive Fine-Tuning teaches models to predict reasoning cost and expected utility before generation, enabling explicit solve-or-skip decisions. Next, Rationality-Aware Reinforcement Learning optimizes sequential decision making under a hard token budget, allowing models to learn long-horizon allocation strategies. Across budgeted mathematical reasoning benchmarks, ROI-Reasoning consistently improves overall score while substantially reducing regret under tight computation budgets.

📄 PDF Abstract BibTeX arXiv:2601.03822

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningMathematical ReasoningDecision Making

Similar Papers 제목 키워드 기반

Rational Metareasoning for Large Language Models

2024-10-07 · C. Nicolò De Sabbata, Theodore R. Sumers, Thomas L. Griffiths

Being prompted to engage in reasoning has emerged as a core technique for using large language models (LLMs), deploying additional inference-time compute to improve task performance. However, as LLMs increase in both siz…

Learning to select computations

2017-11-18 · Frederick Callaway, Sayan Gul, Paul M. Krueger, Thomas L. Griffiths 외

The efficient use of limited computational resources is an essential ingredient of intelligence. Selecting computations optimally according to rational metareasoning would achieve this, but this is computationally intrac…

Management

Ideal Partition of Resources for Metareasoning

2021-10-18 · Eric Horvitz, John Breese

We can achieve significant gains in the value of computation by metareasoning about the nature or extent of base-level problem solving before executing a solution. However, resources that are irrevocably committed to met…

Automated Machine Learning, Bounded Rationality, and Rational Metareasoning

2021-09-10 · Eyke Hüllermeier, Felix Mohr, Alexander Tornede, Marcel Wever

The notion of bounded rationality originated from the insight that perfectly rational behavior cannot be realized by agents with limited cognitive or computational resources. Research on bounded rationality, mainly initi…

AutoMLBIG-bench Machine Learning

Cognitive Load-Aware Inference: A Neuro-Symbolic Framework for Optimizing the Token Economy of Large Language Models

2025-07-01 · Yilun Zhang arxiv

The escalating computational costs of Large Language Model (LLM) inference have become a critical barrier to their widespread and sustainable deployment. While existing optimization strategies are effective, they are pre…

Question AnsweringCode Generation