paper-with-me

홈 › Papers

Distillation Quantification for Large Language Models

2025-01-22 · Sunbowen Lee, Junting Zhou, Chang Ao, Kaige Li, Xinrun Du, Sirui He, Jiaheng Liu, Min Yang, Zhoufutu Wen, Shiwen Ni

Model distillation is a technique for transferring knowledge from large language models (LLMs) to smaller ones, aiming to create resource-efficient yet high-performing models. However, excessive distillation can lead to homogenization, reducing diversity among models and impairing their ability to robustly handle complex or novel tasks. These limitations underscore the need to systematically quantify the distillation process and its impact. In this work, we propose a framework to evaluate and quantify model distillation. Our method addresses two key aspects: (1) Identifying identity cognition contradictions to assess discrepancies in how models perceive and represent identity-related information, and (2) Analyzing multi-granularity response similarities across models to measure the extent of homogenization. Experimental results demonstrate two key insights: (1) Well-known closed-source and open-source LLMs usually exhibit high distillation degrees, except for Claude, Doubao, and Gemini. (2) Base LLMs show higher distillation degrees compared to aligned LLMs. By offering a systematic approach to improve the transparency of LLM data distillation, we call for LLMs with more independent development and more transparent technical reports to improve LLMs' robustness and safety. The code and data are available under https://github.com/Aegis1863/LLMs-Distillation-Quantification.

📄 PDF Abstract BibTeX arXiv:2501.12619

Code (1)

aegis1863/llms-distillation-quantification 공식 구현

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Multi-Teacher Knowledge Distillation via Teacher-Informed Mixture Priors

2026-05-27 · Luyang Fang, Yongkai Chen, Jiazhang Cai, Ping Ma 외 arxiv

Knowledge distillation is a powerful method for model compression, enabling the efficient deployment of complex deep learning models (teachers), including large language models. However, its underlying statistical mechan…

Knowledge DistillationImage ClassificationBayesian InferenceModel Compression

Simple Yet Effective: An Information-Theoretic Approach to Multi-LLM Uncertainty Quantification

2025-07-09 · Maya Kruse, Majid Afshar, Saksham Khatwani, Anoop Mayampurath 외 arxiv

Large language models (LLMs) often behave inconsistently across inputs, indicating uncertainty and motivating the need for its quantification in high-stakes settings. Prior work on calibration and uncertainty quantificat…

Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation

2026-05-07 · Huizi Cui, Huan Ma, Qilin Wang, Yuhang Gao 외 arxiv

Large language models (LLMs) have progressed rapidly in complex reasoning and question answering, yet LLM hallucination remains a central bottleneck that hinders practical deployment, especially for commercial black-box …

Question Answering

Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence

2025-03-18 · Sophia Hager, David Mueller, Kevin Duh, Nicholas Andrews

As large language models (LLMs) are increasingly used for factual question-answering, it becomes more important for LLMs to have the capability to communicate the likelihood that their answer is correct. For these verbal…

Question AnsweringUncertainty Quantification

Knowledge Distillation of Uncertainty using Deep Latent Factor Model

2025-10-22 · Sehyun Park, Jongjin Lee, Yunseop Shin, Ilsang Ohn 외 arxiv

Deep ensembles deliver state-of-the-art, reliable uncertainty quantification, but their heavy computational and memory requirements hinder their practical deployments to real applications such as on-device AI. Knowledge …

Knowledge Distillation