Large Language Model Enhanced Machine Learning Estimators for Classification
Pre-trained large language models (LLM) have emerged as a powerful tool for simulating various scenarios and generating output given specific instructions and multimodal input. In this work, we analyze the specific use of LLM to enhance a classical supervised machine learning method for classification problems. We propose a few approaches to integrate LLM into a classical machine learning estimator to further enhance the prediction performance. We examine the performance of the proposed approaches through both standard supervised learning binary classification tasks, and a transfer learning task where the test data observe distribution changes compared to the training data. Numerical experiments using four publicly available datasets are conducted and suggest that using LLM to enhance classical machine learning estimators can provide significant improvement on prediction performance.
Code (1)
Tasks
Binary ClassificationLanguage ModelingLanguage ModellingLarge Language ModelTransfer LearningSimilar Papers 제목 키워드 기반
Causal Interaction Trees: Tree-Based Subgroup Identification for Observational Data
We propose Causal Interaction Trees for identifying subgroups of participants that have enhanced treatment effects using observational data. We extend the Classification and Regression Tree algorithm by using splitting c…
Decision MakingregressionA Multilevel Approach to Training
We propose a novel training method based on nonlinear multilevel minimization techniques, commonly used for solving discretized large scale partial differential equations. Our multilevel training method constructs a mult…
Optimizing Estimators of Squared Calibration Errors in Classification
In this work, we propose a mean-squared error-based risk that enables the comparison and optimization of estimators of squared calibration errors in practical settings. Improving the calibration of classifiers is crucial…
ClassificationDecision Makingimage-classificationImage Classification+1Are Large Language Models State-of-the-art Quality Estimators for Machine Translation of User-generated Content?
This paper investigates whether large language models (LLMs) are state-of-the-art quality estimators for machine translation of user-generated content (UGC) that contains emotional expressions, without the use of referen…
In-Context LearningMachine Translationparameter-efficient fine-tuningTranslationMachine learning for causal inference: on the use of cross-fit estimators
Modern causal inference methods allow machine learning to be used to weaken parametric modeling assumptions. However, the use of machine learning may result in complications for inference. Doubly-robust cross-fit estimat…
BIG-bench Machine LearningCausal InferenceEnsemble Learning