paper-with-me

홈 › Papers

Compute-Accuracy Pareto Frontiers for Open-Source Reasoning Large Language Models

2025-12-31 · Ákos Prucs, Nara Csutora, Mátyás Antal, Márk Marosi arxiv

Large Language Models (LLMs) are demonstrating rapid improvements on complex reasoning benchmarks, particularly when allowed to utilize intermediate reasoning steps before converging on a final solution. However, current literature often overlooks the significant computational burden associated with generating long reasoning sequences. For industrial applications, model selection depends not only on raw accuracy but also on resource constraints and inference costs. In this work, we conduct a test-time-compute aware evaluation of both contemporary and older open-source LLMs, mapping their Pareto frontiers across math- and reasoning-intensive benchmarks. Our findings identify the Mixture of Experts (MoE) architecture as a strong candidate to balance performance and efficiency in our evaluation setting. Furthermore, we trace the trajectory of Pareto efficiency over time to derive an emergent trend regarding accuracy gain per unit of compute. Finally, we demonstrate that there is a saturation point for inference-time compute. Beyond a certain threshold, accuracy gains diminish, indicating that while extended reasoning capabilities are beneficial, they cannot overcome intrinsic model limitations regarding specific complexities.

📄 PDF Abstract BibTeX arXiv:2512.24776

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The new hybrid COAW method for solving multi-objective problems

2016-01-06 · Zeinab Borhanifar, Elham Shadkam

In this article using Cuckoo Optimization Algorithm and simple additive weighting method the hybrid COAW algorithm is presented to solve multi-objective problems. Cuckoo algorithm is an efficient and structured method fo…

Approximating Pareto Frontiers in Stochastic Multi-Objective Optimization via Hashing and Randomization

2026-04-01 · Jinzhao Li, Nan Jiang, Yexiang Xue arxiv

Stochastic Multi-Objective Optimization (SMOO) is critical for decision-making trading off multiple potentially conflicting objectives in uncertain environments. SMOO aims at identifying the Pareto frontier, which contai…

A Hybrid 2-stage Neural Optimization for Pareto Front Extraction

2021-01-27 · Gurpreet Singh, Soumyajit Gupta, Matthew Lease, Clint Dawson

Classification, recommendation, and ranking problems often involve competing goals with additional constraints (e.g., to satisfy fairness or diversity criteria). Such optimization problems are quite challenging, often in…

Fairness

GaussianPSL: Soft partitioning for complex PSL problem

2025-09-22 · Phuong Mai Dinh, Van-Nam Huynh arxiv

Many practical applications of multi-objective optimization (MOO), including engineering design, autonomous systems, and machine learning, often yield complex Pareto frontiers (e.g., discontinuous, degenerate, or non-con…

Inference economics of language models

2025-06-05 · Ege Erdil

We develop a theoretical model that addresses the economic trade-off between cost per token versus serial token generation speed when deploying LLMs for inference at scale. Our model takes into account arithmetic, memory…