paper-with-me

홈 › Papers

Small Language Model as Data Prospector for Large Language Model

2024-12-13 · Shiwen Ni, Haihong Wu, Di Yang, Qiang Qu, Hamid Alinejad-Rokny, Min Yang

The quality of instruction data directly affects the performance of fine-tuned Large Language Models (LLMs). Previously, \cite{li2023one} proposed \texttt{NUGGETS}, which identifies and selects high-quality quality data from a large dataset by identifying those individual instruction examples that can significantly improve the performance of different tasks after being learnt as one-shot instances. In this work, we propose \texttt{SuperNUGGETS}, an improved variant of \texttt{NUGGETS} optimised for efficiency and performance. Our \texttt{SuperNUGGETS} uses a small language model (SLM) instead of a large language model (LLM) to filter the data for outstanding one-shot instances and refines the predefined set of tests. The experimental results show that the performance of \texttt{SuperNUGGETS} only decreases by 1-2% compared to \texttt{NUGGETS}, but the efficiency can be increased by a factor of 58. Compared to the original \texttt{NUGGETS}, our \texttt{SuperNUGGETS} has a higher utility value due to the significantly lower resource consumption.

📄 PDF Abstract BibTeX arXiv:2412.09990

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language ModelmodelSmall Language Model

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Prospector Heads: Generalized Feature Attribution for Large Models & Data

2024-02-18 · Gautam Machiraju, Alexander Derry, Arjun Desai, Neel Guha 외

Feature attribution, the ability to localize regions of the input data that are relevant for classification, is an important capability for ML models in scientific and biomedical domains. Current methods for feature attr…

Evaluation of Uncertain Inference Models I: PROSPECTOR

2013-03-27 · Robert M. Yadrick, Bruce M. Perrin, David S. Vaughan, Peter D. Holden 외

This paper examines the accuracy of the PROSPECTOR model for uncertain reasoning. PROSPECTOR's solutions for a large number of computer-generated inference networks were compared to those obtained from probability theory…

The Role of Tuning Uncertain Inference Systems

2013-03-27 · Ben P. Wise, Bruce M. Perrin, David S. Vaughan, Robert M. Yadrick

This study examined the effects of "tuning" the parameters of the incremental function of MYCIN, the independent function of PROSPECTOR, a probability model that assumes independence, and a simple additive linear equatio…

Prospector: a mobile app for high-throughput NIRS phenotyping

2021-04-14 · Trevor W. Rife, Chaney Courtney, Jenna Hershberger, Brandon Shaver 외

Quality traits are some of the most important and time-consuming phenotypes to evaluate in plant breeding programs. These traits are often evaluated late in the breeding pipeline due to their cost, resulting in the poten…

Vocal Bursts Intensity Prediction

Comparing Expert Systems Built Using Different Uncertain Inference Systems

2013-03-27 · David S. Vaughan, Bruce M. Perrin, Robert M. Yadrick

This study compares the inherent intuitiveness or usability of the most prominent methods for managing uncertainty in expert systems, including those of EMYCIN, PROSPECTOR, Dempster-Shafer theory, fuzzy set theory, simpl…

regression