paper-with-me

홈 › Papers

Machine Learning Model Attribution Challenge

2023-02-13 · Elizabeth Merkhofer, Deepesh Chaudhari, Hyrum S. Anderson, Keith Manville, Lily Wong, João Gante

We present the findings of the Machine Learning Model Attribution Challenge. Fine-tuned machine learning models may derive from other trained models without obvious attribution characteristics. In this challenge, participants identify the publicly-available base models that underlie a set of anonymous, fine-tuned large language models (LLMs) using only textual output of the models. Contestants aim to correctly attribute the most fine-tuned models, with ties broken in the favor of contestants whose solutions use fewer calls to the fine-tuned models' API. The most successful approaches were manual, as participants observed similarities between model outputs and developed attribution heuristics based on public documentation of the base models, though several teams also submitted automated, statistical solutions.

📄 PDF Abstract BibTeX arXiv:2302.06716

Code (0)

등록된 구현이 없습니다.

Tasks

Attributemodel

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Taming Hyperparameter Sensitivity in Data Attribution: Practical Selection Without Costly Retraining

2025-05-30 · Weiyi Wang, Junwei Deng, Yuzheng Hu, Shiyuan Zhang 외

Data attribution methods, which quantify the influence of individual training data points on a machine learning model, have gained increasing popularity in data-centric applications in modern AI. Despite a recent surge o…

Sensitivity

Feature Attribution from First Principles

2025-05-30 · Magamed Taimeskhanov, Damien Garreau

Feature attribution methods are a popular approach to explain the behavior of machine learning models. They assign importance scores to each input feature, quantifying their influence on the model's prediction. However, …

On the Robustness of Removal-Based Feature Attributions

2023-06-12 · NeurIPS 2023 11

To explain predictions made by complex machine learning models, many feature attribution methods have been developed that assign importance scores to input features. Some recent work challenges the robustness of these me…

Generalized Group Data Attribution

2024-10-13 · Dan Ley, Suraj Srinivas, Shichang Zhang, Gili Rusak 외

Data Attribution (DA) methods quantify the influence of individual training data points on model outputs and have broad applications such as explainability, data selection, and noisy label identification. However, existi…

Computational Efficiency

Greedy PIG: Adaptive Integrated Gradients

2023-11-10 · Kyriakos Axiotis, Sami Abu-al-haija, Lin Chen, Matthew Fahrbach 외

Deep learning has become the standard approach for most machine learning tasks. While its impact is undeniable, interpreting the predictions of deep learning models from a human perspective remains a challenge. In contra…

Deep Learningfeature selection