paper-with-me

홈 › Papers

Are Metrics Enough? Guidelines for Communicating and Visualizing Predictive Models to Subject Matter Experts

2022-05-11 · Ashley Suh, Gabriel Appleby, Erik W. Anderson, Luca Finelli, Remco Chang, Dylan Cashman

Presenting a predictive model's performance is a communication bottleneck that threatens collaborations between data scientists and subject matter experts. Accuracy and error metrics alone fail to tell the whole story of a model - its risks, strengths, and limitations - making it difficult for subject matter experts to feel confident in their decision to use a model. As a result, models may fail in unexpected ways or go entirely unused, as subject matter experts disregard poorly presented models in favor of familiar, yet arguably substandard methods. In this paper, we describe an iterative study conducted with both subject matter experts and data scientists to understand the gaps in communication between these two groups. We find that, while the two groups share common goals of understanding the data and predictions of the model, friction can stem from unfamiliar terms, metrics, and visualizations - limiting the transfer of knowledge to SMEs and discouraging clarifying questions being asked during presentations. Based on our findings, we derive a set of communication guidelines that use visualization as a common medium for communicating the strengths and weaknesses of a model. We provide a demonstration of our guidelines in a regression modeling scenario and elicit feedback on their use from subject matter experts. From our demonstration, subject matter experts were more comfortable discussing a model's performance, more aware of the trade-offs for the presented model, and better equipped to assess the model's risks - ultimately informing and contextualizing the model's use beyond text and numbers.

📄 PDF Abstract BibTeX arXiv:2205.05749

Code (1)

tuftsvalt/modelcomm 공식 구현

Tasks

Friction

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

seeBias: A Comprehensive Tool for Assessing and Visualizing AI Fairness

2025-04-11 · Yilin Ning, Yian Ma, Mingxuan Liu, Xin Li 외

Fairness in artificial intelligence (AI) prediction models is increasingly emphasized to support responsible adoption in high-stakes domains such as health care and criminal justice. Guidelines and implementation framewo…

Fairness

Considerations for Visualizing Uncertainty in Clinical Machine Learning Models

2022-10-21 · Caitlin F. Harrigan, Gabriela Morgenshtern, Anna Goldenberg, Fanny Chevalier

Clinician-facing predictive models are increasingly present in the healthcare setting. Regardless of their success with respect to performance metrics, all models have uncertainty. We investigate how to visually communic…

Deep forecasting of translational impact in medical research

2021-10-17 · Amy PK Nelson, Robert J Gray, James K Ruffle, Henry C Watkins 외

The value of biomedical research--a $1.7 trillion annual investment--is ultimately determined by its downstream, real-world impact. Current objective predictors of impact rest on proxy, reductive metrics of dissemination…

Translation

A Distributed Deep Koopman Learning Algorithm for Control

2024-12-10 · Wenjian Hao, Zehui Lu, Devesh Upadhyay, Shaoshuai Mou

This paper proposes a distributed data-driven framework to address the challenge of dynamics learning from a large amount of training data for optimal control purposes, named distributed deep Koopman learning for control…

Model Predictive Control

Aggregating Concepts of Accuracy and Fairness in Prediction Algorithms

2025-05-13 · David Kinney

An algorithm that outputs predictions about the state of the world will almost always be designed with the implicit or explicit goal of outputting accurate predictions (i.e., predictions that are likely to be true). In a…

Fairness