paper-with-me

Papers

The creative psychometric item generator: a framework for item generation and validation using large language models

2024-08-30 · Antonio Laverghetta Jr., Simone Luchini, Averie Linell, Roni Reiter-Palmon, Roger Beaty

Increasingly, large language models (LLMs) are being used to automate workplace processes requiring a high degree of creativity. While much prior work has examined the creativity of LLMs, there has been little research on whether they can generate valid creativity assessments for humans despite the increasingly central role of creativity in modern economies. We develop a psychometrically inspired framework for creating test items (questions) for a classic free-response creativity test: the creative problem-solving (CPS) task. Our framework, the creative psychometric item generator (CPIG), uses a mixture of LLM-based item generators and evaluators to iteratively develop new prompts for writing CPS items, such that items from later iterations will elicit more creative responses from test takers. We find strong empirical evidence that CPIG generates valid and reliable items and that this effect is not attributable to known biases in the evaluation process. Our findings have implications for employing LLMs to automatically generate valid and reliable creativity tests for humans and AI.

📄 PDF Abstract BibTeX arXiv:2409.00202

Code (0)

등록된 구현이 없습니다.

Tasks

valid

Similar Papers 제목 키워드 기반

Automatic Item Generation for Personality Situational Judgment Tests with Large Language Models

2024-12-10 · Chang-Jin Li, Jiyuan Zhang, Yun Tang, Jian Li

Personality assessment, particularly through situational judgment tests (SJTs), is a vital tool for psychological research, talent selection, and educational evaluation. This study explores the potential of GPT-4, a stat…

Language ModelingLanguage ModellingLarge Language Model

Prompt Engineering for Scale Development in Generative Psychometrics

2026-03-16 · Lara Lee Russell-Lasalandra, Hudson Golino arxiv

This Monte Carlo simulation examines how prompt engineering strategies shape the quality of large language model (LLM)--generated personality assessment items within the AI-GENIE framework for generative psychometrics. I…

Prompt Engineering

Items from Psychometric Tests as Training Data for Personality Profiling Models of Twitter Users

2022-02-21 · WASSA (ACL) 2022 5 · Anne Kreuter, Kai Sassenberg, Roman Klinger

Machine-learned models for author profiling in social media often rely on data acquired via self-reporting-based psychometric tests (questionnaires) filled out by social media users. This is an expensive but accurate dat…

Author ProfilingData Augmentation

The Ultimate Tutorial for AI-driven Scale Development in Generative Psychometrics: Releasing AIGENIE from its Bottle

2026-03-30 · Lara Russell-Lasalandra, Hudson Golino, Luis Eduardo Garrido, Alexander P. Christensen arxiv

Psychological scale development has traditionally required extensive expert involvement, iterative revision, and large-scale pilot testing before psychometric evaluation can begin. The `AIGENIE` R package implements the …

Text Generation

Quantifying Data Contamination in Psychometric Evaluations of LLMs

2025-10-08 · Jongwook Han, Woojung Song, Jonggeun Lee, Yohan Jo arxiv

Recent studies apply psychometric questionnaires to Large Language Models (LLMs) to assess high-level psychological constructs such as values, personality, moral foundations, and dark traits. Although prior work has rais…