paper-with-me

홈 › Papers

Efficient Inference Using Large Language Models with Limited Human Data: Fine-Tuning then Rectification

2025-11-23 · Lei Wang, Zikun Ye, Jinglong Zhao arxiv

Driven by recent advances in artificial intelligence (AI), a growing literature has demonstrated the potential for using large language models (LLMs) as scalable surrogates to generate human-like responses in many business applications. Two common approaches to improve the performance of LLMs include: fine-tuning, which aligns LLMs more closely with human responses, and rectification, which corrects biases in LLM outputs. In this paper, we develop a two-stage framework that combines fine-tuning and rectification, and optimally allocates limited labeled samples across the two stages. Unlike the conventional objective that minimizes the mean squared prediction errors, we propose to minimize the variance of the prediction errors as the fine-tuning objective, which is optimal for the downstream rectification stage. Building on this insight, we leverage the scaling law of fine-tuning to optimally allocate the limited labeled human data between the fine-tuning and rectification stages. Our empirical analysis validates the fine-tuning scaling law and confirms that our proposed optimal allocation rule reliably identifies the optimal sample allocation. We demonstrate substantial efficiency gains in estimation and inference performance relative to fine-tuning or rectification alone, or to employing the standard mean-squared error objective within the fine-tuning then rectification framework, resulting in significant cost savings for reliable business decisions.

📄 PDF Abstract BibTeX arXiv:2511.19486

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Perceptions to Beliefs: Exploring Precursory Inferences for Theory of Mind in Large Language Models

2024-07-08 · Chani Jung, Dongkwan Kim, Jiho Jin, Jiseon Kim 외

While humans naturally develop theory of mind (ToM), the capability to understand other people's mental states and beliefs, state-of-the-art large language models (LLMs) underperform on simple ToM benchmarks. We posit th…

Can Large Language Models Capture Dissenting Human Voices?

2023-05-23 · Noah Lee, Na Min An, James Thorne

Large language models (LLMs) have shown impressive achievements in solving a broad range of tasks. Augmented by instruction fine-tuning, LLMs have also been shown to generalize in zero-shot settings as well. However, whe…

Natural Language InferenceNatural Language Understanding

A large annotated corpus for learning natural language inference

2015-08-21 · EMNLP 2015 9 · Samuel R. Bowman, Gabor Angeli, Christopher Potts, Christopher D. Manning

Understanding entailment and contradiction is fundamental to understanding natural language, and inference about entailment and contradiction is a valuable testing ground for the development of semantic representations. …

Image CaptioningNatural Language InferenceSentence

Comparing Human and Large Language Model Interpretation of Implicit Information

2026-04-18 · Antonio De Santis, Tommaso Bonetti, Andrea Tocchetti, Marco Brambilla arxiv

The interpretation of implicit meanings is an integral aspect of human communication. However, this framework may not transfer to interactions with Large Language Models (LLMs). To investigate this, we introduce the task…

Information Extraction

OCNLI: Original Chinese Natural Language Inference

2020-10-12 · Findings of the Association for Computational Linguistics 2020 · Hai Hu, Kyle Richardson, Liang Xu, Lu Li 외

Despite the tremendous recent progress on natural language inference (NLI), driven largely by large-scale investment in new datasets (e.g., SNLI, MNLI) and advances in modeling, most progress has been limited to English …

Natural Language InferenceSentenceTranslation