paper-with-me

홈 › Papers

On the Utility of Domain-Adjacent Fine-Tuned Model Ensembles for Few-shot Problems

2024-06-19 · Md Ibrahim Ibne Alam, Parikshit Ram, Soham Dan, Horst Samulowitz, Koushik Kar

Large Language Models (LLMs) have been observed to perform well on a wide range of downstream tasks when fine-tuned on domain-specific data. However, such data may not be readily available in many applications, motivating zero-shot or few-shot approaches using domain-adjacent models. While several fine-tuned models for various tasks are available, finding an appropriate domain-adjacent model for a given task is often not straight forward. In this paper, we study DAFT-E, a framework that utilizes an Ensemble of Domain-Adjacent Fine-Tuned Foundation Models for few-shot problems. We show that for zero-shot problems, this ensembling method provides an accuracy performance close to that of the single best model. With few-shot problems, this performance improves further, at which point DEFT-E can outperform any single domain-adjacent model while requiring much less data for domain-specific fine-tuning.

📄 PDF Abstract BibTeX arXiv:2406.13720

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Uncertainty quantification in fine-tuned LLMs using LoRA ensembles

2024-02-19 · Oleksandr Balabanov, Hampus Linander

Fine-tuning large language models can improve task specific performance, although a general understanding of what the fine-tuned model has learned, forgotten and how to trust its predictions is still missing. We derive p…

Multiple-choiceUncertainty Quantification

Small or Large? Zero-Shot or Finetuned? Guiding Language Model Choice for Specialized Applications in Healthcare

2025-04-29 · Lovedeep Gondara, Jonathan Simkin, Graham Sayle, Shebnum Devji 외

This study aims to guide language model selection by investigating: 1) the necessity of finetuning versus zero-shot usage, 2) the benefits of domain-adjacent versus generic pretrained models, 3) the value of further doma…

Language ModelingLanguage ModellingModel Selection

The University of Edinburgh’s Bengali-Hindi Submissions to the WMT21 News Translation Task

2021-11-01 · WMT (EMNLP) 2021 11 · Proyag Pal, Alham Fikri Aji, Pinzhen Chen, Sukanta Sen

We describe the University of Edinburgh’s Bengali\leftrightarrowHindi constrained systems submitted to the WMT21 News Translation task. We submitted ensembles of Transformer models built with large-scale back-translation…

Translation

Ensembling Finetuned Language Models for Text Classification

2024-10-25 · Sebastian Pineda Arango, Maciej Janowski, Lennart Purucker, Arber Zela 외

Finetuning is a common practice widespread across different communities to adapt pretrained models to particular tasks. Text classification is one of these tasks for which many pretrained models are available. On the oth…

Classificationtext-classificationText Classification

COVID-19 Classification Using Staked Ensembles: A Comprehensive Analysis

2020-10-07 · Lalith Bharadwaj B, Rohit Boddeda, Sai Vardhan K, Madhu G

The issue of COVID-19, increasing with a massive mortality rate. This led to the WHO declaring it as a pandemic. In this situation, it is crucial to perform efficient and fast diagnosis. The reverse transcript polymerase…

Classification