paper-with-me

Papers

Should Machine Learning Models Report to Us When They Are Clueless?

2022-03-23 · Roozbeh Yousefzadeh, Xuenan Cao

The right to AI explainability has consolidated as a consensus in the research community and policy-making. However, a key component of explainability has been missing: extrapolation, which describes the extent to which AI models can be clueless when they encounter unfamiliar samples (i.e., samples outside the convex hull of their training sets, as we will explain). We report that AI models extrapolate outside their range of familiar data, frequently and without notifying the users and stakeholders. Knowing whether a model has extrapolated or not is a fundamental insight that should be included in explaining AI models in favor of transparency and accountability. Instead of dwelling on the negatives, we offer ways to clear the roadblocks in promoting AI transparency. Our analysis commentary accompanying practical clauses useful to include in AI regulations such as the National AI Initiative Act in the US and the AI Act by the European Commission.

📄 PDF Abstract BibTeX arXiv:2203.12131

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Are Prompt-based Models Clueless?

2022-05-19 · ACL 2022 5 · Pride Kavumba, Ryo Takahashi, Yusuke Oda

Finetuning large pre-trained language models with a task-specific head has advanced the state-of-the-art on many natural language understanding benchmarks. However, models with a task-specific head require a lot of train…

Language ModelingLanguage ModellingNatural Language Understanding

RadFlag: A Black-Box Hallucination Detection Method for Medical Vision Language Models

2024-11-01 · Serena Zhang, Sraavya Sambara, Oishi Banerjee, Julian Acosta 외

Generating accurate radiology reports from medical images is a clinically important but challenging task. While current Vision Language Models (VLMs) show promise, they are prone to generating hallucinations, potentially…

HallucinationLanguage ModelingLanguage ModellingLarge Language Model

Language model developers should report train-test overlap

2024-10-10 · Andy K Zhang, Kevin Klyman, Yifan Mai, Yoav Levine 외

Language models are extensively evaluated, but correctly interpreting evaluation results requires knowledge of train-test overlap which refers to the extent to which the language model is trained on the very data it is b…

Language ModelingLanguage Modellingmodel

A look under the hood of the Interactive Deep Learning Enterprise (No-IDLE)

2024-06-27 · Daniel Sonntag, Michael Barz, Thiago Gouvêa

This DFKI technical report presents the anatomy of the No-IDLE prototype system (funded by the German Federal Ministry of Education and Research) that provides not only basic and fundamental research in interactive machi…

AnatomyDeep Learningmultimodal interaction

Conscious Intelligence Requires Lifelong Autonomous Programming For General Purposes

2020-06-30

Universal Turing Machines [29, 10, 18] are well known in computer science but they are about manual programming for general purposes. Although human children perform conscious learning (i.e., learning while being conscio…

Natural Language Understanding