paper-with-me

홈 › Papers

On Predicting the Post-training Potential of Pre-trained LLMs

2026-05-12 · Xiaoyuan Li, Yubo Ma, Kexin Yang, Moxin Li, Keqin Bao, Wenie Wang, Fuli Feng, Dayiheng Liu arxiv

The performance of Large Language Models (LLMs) on downstream tasks is fundamentally constrained by the capabilities acquired during pre-training. However, traditional benchmarks like MMLU often fail to reflect a base model's plasticity in complex open-ended scenarios, leading to inefficient model selection. We address this by introducing a new task of predicting post-training potential - forecasting a base model's performance before post-training. We propose RuDE (Rubric-based Discriminative Evaluation), a unified framework that bypasses the generation gap of base models by leveraging response discrimination. Guided by our systematic 4C Taxonomy, RuDE constructs controlled contrastive pairs across diverse domains by fine-grained rubric violations. Extensive experiments demonstrate a correlation greater than 90% with post-training performance. Crucially, validation via Reinforcement Learning (RL) confirms that RuDE effectively identifies high-potential smaller models that outperform larger counterparts, offering a compute-efficient mechanism for foundation model development.

📄 PDF Abstract BibTeX arXiv:2605.11978

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

The Foundational Capabilities of Large Language Models in Predicting Postoperative Risks Using Clinical Notes

2024-02-27 · Charles Alba, Bing Xue, Joanna Abraham, Thomas Kannampallil 외

Clinical notes recorded during a patient's perioperative journey holds immense informational value. Advances in large language models (LLMs) offer opportunities for bridging this gap. Using 84,875 pre-operative notes and…

Domain AdaptationMulti-Task LearningWord Embeddings

How Post-Training Reshapes LLMs: A Mechanistic View on Knowledge, Truthfulness, Refusal, and Confidence

2025-04-03 · Hongzhe Du, Weikai Li, Min Cai, Karim Saraipour 외

Post-training is essential for the success of large language models (LLMs), transforming pre-trained base models into more useful and aligned post-trained models. While plenty of works have studied post-training algorith…

Crowdsourcing with Enhanced Data Quality Assurance: An Efficient Approach to Mitigate Resource Scarcity Challenges in Training Large Language Models for Healthcare

2024-05-16 · P. Barai, G. Leroy, P. Bisht, J. M. Rothman 외

Large Language Models (LLMs) have demonstrated immense potential in artificial intelligence across various domains, including healthcare. However, their efficacy is hindered by the need for high-quality labeled data, whi…

Decision Making

Finding and Reactivating Post-Trained LLMs' Hidden Safety Mechanisms

2026-03-10 · Mingjie Li, Wai Man Si, Michael Backes, Yang Zhang 외 arxiv

Despite the impressive performance of general-purpose large language models (LLMs), they often require fine-tuning or post-training to excel at specific tasks. For instance, large reasoning models (LRMs), such as the Dee…

Large language models surpass human experts in predicting neuroscience results

2024-03-04 · Xiaoliang Luo, Akilles Rechardt, Guangzhi Sun, Kevin K. Nejad 외

Scientific discoveries often hinge on synthesizing decades of research, a task that potentially outstrips human information processing capacities. Large language models (LLMs) offer a solution. LLMs trained on the vast s…