paper-with-me

홈 › Papers

What can AI do for me: Evaluating Machine Learning Interpretations in Cooperative Play

2018-10-23 · Shi Feng, Jordan Boyd-Graber

Machine learning is an important tool for decision making, but its ethical and responsible application requires rigorous vetting of its interpretability and utility: an understudied problem, particularly for natural language processing models. We propose an evaluation of interpretation on a real task with real human users, where the effectiveness of interpretation is measured by how much it improves human performance. We design a grounded, realistic human-computer cooperative setting using a question answering task, Quizbowl. We recruit both trivia experts and novices to play this game with computer as their teammate, who communicates its prediction via three different interpretations. We also provide design guidance for natural language processing human-in-the-loop settings.

📄 PDF Abstract BibTeX arXiv:1810.09648

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningDecision MakingQuestion Answering

Methods 이 논문이 사용한 방법론

Interpretability 설명 없음

Similar Papers 제목 키워드 기반

A mathematical theory of cooperative communication

2019-10-07 · NeurIPS 2020 12 · Pei Wang, Junqi Wang, Pushpi Paranamana, Patrick Shafto

Cooperative communication plays a central role in theories of human cognition, language, development, culture, and human-robot interaction. Prior models of cooperative communication are algorithmic in nature and do not s…

Cultural Vocal Bursts Intensity Prediction

Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game

2026-04-22 · Saad Mankarious arxiv

We introduce \emph{Quantum Frog}, a two-player cooperative game built on a novel \emph{quantized-time} mechanic in which the environment advances only when a player acts. Inspired by the classic arcade game Frogger, Quan…

Reinforcement Learning

Computational Hermeneutics: Evaluating generative AI as a cultural technology

2026-03-31 · Cody Kommers, Ruth Ahnert, Maria Antoniak, Emmanouil Benetos 외 arxiv

Generative AI systems are increasingly recognized as cultural technologies, yet current evaluation frameworks often treat culture as a variable to be measured rather than fundamental to the system's operation. Drawing on…

Any-Play: An Intrinsic Augmentation for Zero-Shot Coordination

2022-01-28 · Keane Lucas, Ross E. Allen

Cooperative artificial intelligence with human or superhuman proficiency in collaborative tasks stands at the frontier of machine learning research. Prior work has tended to evaluate cooperative AI performance under the …

DecoderDiversity

Generalizing the theory of cooperative inference

2018-10-04 · Pei Wang, Pushpi Paranamana, Patrick Shafto

Cooperation information sharing is important to theories of human learning and has potential implications for machine learning. Prior work derived conditions for achieving optimal Cooperative Inference given strong, rela…

BIG-bench Machine Learning