paper-with-me

홈 › Papers

Predicting Human Psychometric Properties Using Computational Language Models

2022-05-12 · Antonio Laverghetta Jr., Animesh Nighojkar, Jamshidbek Mirzakhalov, John Licato

Transformer-based language models (LMs) continue to achieve state-of-the-art performance on natural language processing (NLP) benchmarks, including tasks designed to mimic human-inspired "commonsense" competencies. To better understand the degree to which LMs can be said to have certain linguistic reasoning skills, researchers are beginning to adapt the tools and concepts from psychometrics. But to what extent can benefits flow in the other direction? In other words, can LMs be of use in predicting the psychometric properties of test items, when those items are given to human participants? If so, the benefit for psychometric practitioners is enormous, as it can reduce the need for multiple rounds of empirical testing. We gather responses from numerous human participants and LMs (transformer- and non-transformer-based) on a broad diagnostic test of linguistic competencies. We then use the human responses to calculate standard psychometric properties of the items in the diagnostic test, using the human responses and the LM responses separately. We then determine how well these two sets of predictions correlate. We find that transformer-based LMs predict the human psychometric data consistently well across most categories, suggesting that they can be used to gather human-like psychometric data without the need for extensive human trials.

📄 PDF Abstract BibTeX arXiv:2205.06203

Code (0)

등록된 구현이 없습니다.

Tasks

Diagnostic

Similar Papers 제목 키워드 기반

Can Transformer Language Models Predict Psychometric Properties?

2021-06-12 · Joint Conference on Lexical and Computational Semantics 2021 · Antonio Laverghetta Jr., Animesh Nighojkar, Jamshidbek Mirzakhalov, John Licato

Transformer-based language models (LMs) continue to advance state-of-the-art performance on NLP benchmark tasks, including tasks designed to mimic human-inspired "commonsense" competencies. To better understand the degre…

Diagnostic

AIPsychoBench: Understanding the Psychometric Differences between LLMs and Humans

2025-09-20 · Wei Xie, Shuoyoucheng Ma, Zhenhua Wang, Enze Wang 외 arxiv

Large Language Models (LLMs) with hundreds of billions of parameters have exhibited human-like intelligence by learning from vast amounts of internet-scale data. However, the uninterpretability of large-scale neural netw…

Improving LLM Leaderboards with Psychometrical Methodology

2025-01-27 · Denis Federiakin

The rapid development of large language models (LLMs) has necessitated the creation of benchmarks to evaluate their performance. These benchmarks resemble human tests and surveys, as they consist of sets of questions des…

Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis

2024-05-12 · Nikolay B Petrov, Gregory Serapio-García, Jason Rentfrow

The humanlike responses of large language models (LLMs) have prompted social scientists to investigate whether LLMs can be used to simulate human participants in experiments, opinion polls and surveys. Of central interes…

Multiple-choiceQuestion Answering

Human Psychometric Questionnaires Mischaracterize LLM Behavior

2025-09-12 · Woojung Song, Dongmin Choi, Yoonah Park, Jongwook Han 외 arxiv

We examine whether human psychometric questionnaires can serve as reliable tools for characterizing and predicting LLM behavior in everyday user interactions. We analyze eight open-source LLMs by comparing their value an…