paper-with-me

홈 › Papers

AutoFT: Learning an Objective for Robust Fine-Tuning

2024-01-18 · Caroline Choi, Yoonho Lee, Annie Chen, Allan Zhou, aditi raghunathan, Chelsea Finn

Foundation models encode rich representations that can be adapted to downstream tasks by fine-tuning. However, fine-tuning a model on one data distribution often degrades performance under distribution shifts. Current approaches to robust fine-tuning use hand-crafted regularization techniques to constrain the fine-tuning process towards the pretrained model. Yet, it is hard to specify how to adapt relevant characteristics of the foundation model during fine-tuning, as this depends on how the pre-training, fine-tuning, and test data distributions relate to each other. We propose AutoFT, a data-driven approach for robust fine-tuning. Given a task, AutoFT searches for a fine-tuning procedure that enhances out-of-distribution (OOD) generalization. Specifically, AutoFT uses bi-level optimization to search for an objective function and hyperparameters that maximize post-adaptation performance on a small OOD validation set. We evaluate AutoFT on nine natural distribution shifts. Our experiments show that AutoFT significantly improves generalization to OOD inputs, outperforming existing robust fine-tuning methods. Notably, AutoFT achieves a new state-of-the-art on the WILDS iWildCam and FMoW benchmarks, outperforming the previous best methods by $6.0\%$ and $1.5\%$, respectively.

📄 PDF Abstract BibTeX arXiv:2401.10220

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Weight Decay 설명 없음
BASE 설명 없음

Similar Papers 제목 키워드 기반

AutoFT: Automatic Fine-Tune for Parameters Transfer Learning in Click-Through Rate Prediction

2021-06-09 · Xiangli Yang, Qing Liu, Rong Su, Ruiming Tang 외

Recommender systems are often asked to serve multiple recommendation scenarios or domains. Fine-tuning a pre-trained CTR model from source domains and adapting it to a target domain allows knowledge transferring. However…

Click-Through Rate PredictionRecommendation SystemsTransfer Learning

Objective Matters: Fine-Tuning Objectives Shape Safety, Robustness, and Persona Drift

2026-01-19 · Daniel Vennemeyer, Punya Syon Pandey, Phan Anh Duong, Michael Umeokoli 외 arxiv

Fine-tuning LLMs on benign data can still degrade alignment and adversarial robustness, yet direct analysis of the role of fine-tuning objectives in shaping these safety outcomes remain limited. We present a controlled c…

Adversarial Robustness

Aligning the Pretraining and Finetuning Objectives of Language Models

2020-02-05 · Nuo Wang Pierse, Jingwen Lu

We demonstrate that explicitly aligning the pretraining objectives to the finetuning objectives in language model training significantly improves the finetuning task performance and reduces the minimum amount of finetuni…

Language ModelingLanguage Modelling

Conditional Language Policy: A General Framework for Steerable Multi-Objective Finetuning

2024-07-22 · Kaiwen Wang, Rahul Kidambi, Ryan Sullivan, Alekh Agarwal 외

Reward-based finetuning is crucial for aligning language policies with intended behaviors (e.g., creativity and safety). A key challenge is to develop steerable language models that trade-off multiple (conflicting) objec…

Declaration-based Prompt Tuning for Visual Question Answering

2022-05-05 · Yuhang Liu, Wei Wei, Daowan Peng, Feida Zhu

In recent years, the pre-training-then-fine-tuning paradigm has yielded immense success on a wide spectrum of cross-modal tasks, such as visual question answering (VQA), in which a visual-language (VL) model is first opt…

Image-text matchingLanguage ModelingLanguage ModellingMasked Language Modeling+5