paper-with-me

홈 › Papers

Can vectors read minds better than experts? Comparing data augmentation strategies for the automated scoring of children's mindreading ability

2021-06-03 · ACL 2021 5 · Venelin Kovatchev, Phillip Smith, Mark Lee, Rory Devine

In this paper we implement and compare 7 different data augmentation strategies for the task of automatic scoring of children's ability to understand others' thoughts, feelings, and desires (or "mindreading"). We recruit in-domain experts to re-annotate augmented samples and determine to what extent each strategy preserves the original rating. We also carry out multiple experiments to measure how much each augmentation strategy improves the performance of automatic scoring systems. To determine the capabilities of automatic systems to generalize to unseen data, we create UK-MIND-20 - a new corpus of children's performance on tests of mindreading, consisting of 10,320 question-answer pairs. We obtain a new state-of-the-art performance on the MIND-CA corpus, improving macro-F1-score by 6 points. Results indicate that both the number of training examples and the quality of the augmentation strategies affect the performance of the systems. The task-specific augmentations generally outperform task-agnostic augmentations. Automatic augmentations based on vectors (GloVe, FastText) perform the worst. We find that systems trained on MIND-CA generalize well to UK-MIND-20. We demonstrate that data augmentation strategies also improve the performance on unseen data.

📄 PDF Abstract BibTeX arXiv:2106.01635

Code (0)

등록된 구현이 없습니다.

Tasks

Data Augmentation

Similar Papers 제목 키워드 기반

Cognitive networks highlight differences and similarities in the STEM mindsets of human and LLM-simulated trainees, experts and academics

2025-02-26 · Edith Haim, Lars van den Bergh, Cynthia S. Q. Siew, Yoed N. Kenett 외

Understanding attitudes towards STEM means quantifying the cognitive and emotional ways in which individuals, and potentially large language models too, conceptualise such subjects. This study uses behavioural forma ment…

Clustering

The Reader is the Metric: How Textual Features and Reader Profiles Explain Conflicting Evaluations of AI Creative Writing

2025-06-03 · Guillermo Marco, Julio Gonzalo, Víctor Fresno

Recent studies comparing AI-generated and human-authored literary texts have produced conflicting results: some suggest AI already surpasses human quality, while others argue it still falls short. We start from the hypot…

Feature ImportanceSentenceText Generation

MindSearch: Mimicking Human Minds Elicits Deep AI Searcher

2024-07-29 · Zehui Chen, Kuikun Liu, Qiuchen Wang, Jiangning Liu 외

Information seeking and integration is a complex cognitive task that consumes enormous time and effort. Inspired by the remarkable progress of Large Language Models, recent works attempt to solve this task by combining L…

2D Semantic Segmentation task 1 (8 classes)graph constructionInformation Retrieval

Mindstorms in Natural Language-Based Societies of Mind

2023-05-26 · Mingchen Zhuge, Haozhe Liu, Francesco Faccio, Dylan R. Ashley 외

Both Minsky's "society of mind" and Schmidhuber's "learning to think" inspire diverse societies of large multimodal neural networks (NNs) that solve problems by interviewing each other in a "mindstorm." Recent implementa…

3D GenerationImage CaptioningImage GenerationQuestion Answering+1

Supporting decisions by unleashing multiple mindsets using pairwise comparisons method

2021-07-04 · Salvatore Greco, Sajid Siraj, Michele Lundy

Inconsistency in pairwise comparison judgements is often perceived as an unwanted phenomenon and researchers have proposed a number of techniques to either reduce it or to correct it. We take a viewpoint that this incons…