paper-with-me

홈 › Papers

Coupled Calibration and Learning: Mitigating Teacher Bias in LLM Distillation without Target-Domain Reward Feedback

2026-09-15 · Haichen Hu, Yuheng Zhang, David Simchi-Levi arxiv

Large language model (LLM) distillation aims to transfer the capabilities of a powerful teacher to a smaller student. Direct imitation, however, can also transfer the teacher's systematic bias and errors. This challenge is particularly pronounced under covariate shift, when the teacher's reliability on target questions is uncertain and target-domain reward feedback is unavailable. We propose Coupled Calibration and Learning (CCL), an LLM distillation algorithm that couples teacher calibration with student updates through token-level branching, using reward feedback only on source questions. Each iteration calibrates the teacher using source feedback and then uses the calibrated teacher to train the student on target questions. The updated student, in turn, informs subsequent calibration. In an autoregressive policy framework, we prove that the output student's expected average Kullback-Leibler divergence to the oracle student converges to zero at a polynomial rate in the number of iterations. The oracle maximizes the true reference-regularized target reward within the student class, which need not represent the unrestricted optimal policy. Our analysis quantifies the progress of projected student gradient updates while controlling the error in teacher calibration. We further establish a separation from regularized direct matching: its error relative to the oracle student can remain bounded away from zero even when the teacher achieves higher regularized target reward than every student policy. These results demonstrate that LLM distillation can overcome persistent teacher bias and recover the optimal student through coupled calibration and learning, without target-domain reward feedback.

📄 PDF Abstract BibTeX arXiv:2609.17474

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Dual-Teacher De-biasing Distillation Framework for Multi-domain Fake News Detection

2023-12-02 · Jiayang Li, Xuan Feng, Tianlong Gu, Liang Chang

Multi-domain fake news detection aims to identify whether various news from different domains is real or fake and has become urgent and important. However, existing methods are dedicated to improving the overall performa…

Fake News DetectionKnowledge Distillation

Adaptive Decoupled Pose Knowledge Distillation

2023-10-01 · journal 2023 10 · Jie Xu, Shanshan Zhang, and Jian Yang

Existing state-of-the-art human pose estimation approaches require heavy computational resources for accurate prediction. One promising technique to obtain an accurate yet lightweight pose estimator is Knowledge Distilla…

Knowledge DistillationPose Estimation

Teacher's pet: understanding and mitigating biases in distillation

2021-06-19 · Michal Lukasik, Srinadh Bhojanapalli, Aditya Krishna Menon, Sanjiv Kumar

Knowledge distillation is widely used as a means of improving the performance of a relatively simple student model using the predictions from a complex teacher model. Several works have shown that distillation significan…

image-classificationImage ClassificationKnowledge Distillation

Reliability Gated Multi-Teacher Distillation for Low Resource Abstractive Summarization

2026-04-03 · Dipto Sumit, Ankan Kumar Roy, Sadia Khair Rodela, Atia Haque Asha 외 arxiv

We study multiteacher knowledge distillation for low resource abstractive summarization from a reliability aware perspective. We introduce EWAD (Entropy Weighted Agreement Aware Distillation), a token level mechanism tha…

Knowledge DistillationSemantic Similarity

Backtracking When It Strays: Mitigating Dual Exposure Biases in LLM Reasoning Distillation

2026-05-19 · Bing Wang, Shaotian Yan, Chen Shen, kaiyuan liu 외 arxiv

Large language models (LLMs) have achieved remarkable success in complex reasoning tasks via long chain-of-thought (CoT), yet their immense computational overhead hinders real-world deployment. LLM reasoning distillation…