paper-with-me

홈 › Papers

Capturing Humans' Mental Models of AI: An Item Response Theory Approach

2023-05-15 · Markelle Kelly, Aakriti Kumar, Padhraic Smyth, Mark Steyvers

Improving our understanding of how humans perceive AI teammates is an important foundation for our general understanding of human-AI teams. Extending relevant work from cognitive science, we propose a framework based on item response theory for modeling these perceptions. We apply this framework to real-world experiments, in which each participant works alongside another person or an AI agent in a question-answering setting, repeatedly assessing their teammate's performance. Using this experimental data, we demonstrate the use of our framework for testing research questions about people's perceptions of both AI agents and other people. We contrast mental models of AI teammates with those of human teammates as we characterize the dimensionality of these mental models, their development over time, and the influence of the participants' own self-perception. Our results indicate that people expect AI agents' performance to be significantly better on average than the performance of other humans, with less variation across different types of problems. We conclude with a discussion of the implications of these findings for human-AI interaction.

📄 PDF Abstract BibTeX arXiv:2305.09064

Code (1)

markellekelly/ai_mental_models 공식 구현

Tasks

AI AgentQuestion Answering

Similar Papers 제목 키워드 기반

Amortised Design Optimization for Item Response Theory

2023-07-19 · Antti Keurulainen, Isak Westerlund, Oskar Keurulainen, Andrew Howes

Item Response Theory (IRT) is a well known method for assessing responses from humans in education and psychology. In education, IRT is used to infer student abilities and characteristics of test items from student respo…

Deep Reinforcement LearningExperimental Design

Psychometric Alignment: Capturing Human Knowledge Distributions via Language Models

2024-07-22 · Joy He-Yueya, Wanjing Anya Ma, Kanishk Gandhi, Benjamin W. Domingue 외

Language models (LMs) are increasingly used to simulate human-like responses in scenarios where accurately mimicking a population's behavior can guide decision-making, such as in developing educational materials and desi…

Variational Item Response Theory: Fast, Accurate, and Expressive

2020-02-01 · Mike Wu, Richard L. Davis, Benjamin W. Domingue, Chris Piech 외

Item Response Theory (IRT) is a ubiquitous model for understanding humans based on their responses to questions, used in fields as diverse as education, medicine and psychology. Large modern datasets offer opportunities …

Bayesian Inference

ALBA: Adaptive Language-based Assessments for Mental Health

2023-11-11 · Vasudha Varadarajan, Sverker Sikström, Oscar N. E. Kjell, H. Andrew Schwartz

Mental health issues differ widely among individuals, with varied signs and symptoms. Recently, language-based assessments have shown promise in capturing this diversity, but they require a substantial sample of words pe…

Diversity

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory

2026-01-31 · Junhyuk Choi, Sohhyung Park, Chanhee Cho, Hyeonchu Park 외 arxiv

While LLM-as-a-Judge is widely used in automated evaluation, existing validation practices primarily operate at the level of observed outputs, offering limited insight into whether LLM judges themselves function as stabl…