paper-with-me

Papers

The Effect of Idea Elaboration on the Automatic Assessment of Idea Originality

2026-04-22 · Umberto Domanti, Moritz Mock, Sergio Agnoli, Antonella De Angeli arxiv

Automatic systems are increasingly used to assess the originality of responses in creative tasks. They offer a potential solution to key limitations of human assessment (cost, fatigue, and subjectivity), but there is preliminary evidence of a self-preference bias. Accordingly, automatic systems tend to prefer outcomes that are more closely related to their style, rather than to the human one. In this paper, we investigated how Large Language Models (LLMs) align with human raters in assessing the originality of responses in a divergent thinking task. We analysed 4,813 responses to the Alternate Uses Task produced by higher and lower creative humans and ChatGPT-4o. Human raters were two university students who underwent intensive training. Machine raters were two specialised systems fine-tuned on AUT responses and corresponding human ratings (OCSAI and CLAUS) and ChatGPT-4o, which was prompted with the same instructions as human raters. Results confirmed the presence of a self-preference bias in LLMs. Automatic systems tended to privilege artificial responses. However, this self-preference bias disappeared when the analyses controlled for the idea elaboration. We discuss theoretical and methodological implications of these findings by highlighting future directions for research on creativity assessment.

📄 PDF Abstract BibTeX arXiv:2604.20569

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A type-theoretical approach to Universal Grammar

2015-05-16 · Erkki Luuk

The idea of Universal Grammar (UG) as the hypothetical linguistic structure shared by all human languages harkens back at least to the 13th century. The best known modern elaborations of the idea are due to Chomsky. Foll…

Vocal Bursts Type Prediction

Another Turn, Better Output? A Turn-Wise Analysis of Iterative LLM Prompting

2025-09-08 · Shashidhar Reddy Javaji, Bhavul Gauri, Zining Zhu arxiv

Large language models (LLMs) are now used in multi-turn workflows, but we still lack a clear way to measure when iteration helps and when it hurts. We present an evaluation framework for iterative refinement that spans i…

AI Research Agents Narrow Scientific Exploration

2026-05-27 · Yixuan Tang, Yi Yang arxiv

AI research agents now support large-scale AI-assisted scientific discovery. We examine whether AI-generated ideas broaden scientific exploration or primarily reinforce existing work. Using five agent frameworks and five…

Automatic Assessment of Divergent Thinking in Chinese Language with TransDis: A Transformer-Based Language Model Approach

2023-06-26 · Tianchen Yang, Qifan Zhang, Zhaoyang Sun, Yubo Hou

Language models have been increasingly popular for automatic creativity assessment, generating semantic distances to objectively measure the quality of creative ideas. However, there is currently a lack of an automatic a…

Language ModelingLanguage Modellingvalid

The AI Memory Gap: Users Misremember What They Created With AI or Without

2025-09-15 · Tim Zindulka, Sven Goller, Daniela Fernandes, Robin Welsch 외 arxiv

As large language models (LLMs) become embedded in interactive text generation, disclosure of AI as a source depends on people remembering which ideas or texts came from themselves and which were created with AI. We inve…

Text Generation