paper-with-me

Papers

Work Hard, Play Hard: Collecting Acceptability Annotations through a 3D Game

2022-06-01 · LREC 2022 6 · Federico Bonetti, Elisa Leonardelli, Daniela Trotta, Raffaele Guarasci, Sara Tonelli

Corpus-based studies on acceptability judgements have always stimulated the interest of researchers, both in theoretical and computational fields. Some approaches focused on spontaneous judgements collected through different types of tasks, others on data annotated through crowd-sourcing platforms, still others relied on expert annotated data available from the literature. The release of CoLA corpus, a large-scale corpus of sentences extracted from linguistic handbooks as examples of acceptable/non acceptable phenomena in English, has revived interest in the reliability of judgements of linguistic experts vs. non-experts. Several issues are still open. In this work, we contribute to this debate by presenting a 3D video game that was used to collect acceptability judgments on Italian sentences. We analyse the resulting annotations in terms of agreement among players and by comparing them with experts’ acceptability judgments. We also discuss different game settings to assess their impact on participants’ motivation and engagement. The final dataset containing 1,062 sentences, which were selected based on majority voting, is released for future research and comparisons.

📄 PDF Abstract BibTeX

Code (1)

dhfbk/itacola-dataset 공식 구현

Tasks

CoLA

Similar Papers 제목 키워드 기반

Grammatical Analysis of Pretrained Sentence Encoders with Acceptability Judgments

2018-12-11 · Anonymous

Recent pretrained sentence encoders achieve state of the art results on language understanding tasks, but does this mean they have implicit knowledge of syntactic structures? We introduce a grammatically annotated develo…

CoLALinguistic AcceptabilitySentence

‘Meet me at the ribary’ – Acceptability of spelling variants in free-text answers to listening comprehension prompts

2022-07-01 · NAACL (BEA) 2022 7 · Ronja Laarmann-Quante, Leska Schwarz, Andrea Horbach, Torsten Zesch

When listening comprehension is tested as a free-text production task, a challenge for scoring the answers is the resulting wide range of spelling variants. When judging whether a variant is acceptable or not, human rate…

The Acceptability Delta Criterion: Testing Knowledge of Language using the Gradience of Sentence Acceptability

2021-11-01 · EMNLP (BlackboxNLP) 2021 11 · Héctor Vázquez Martínez

Any test that promises to assess Human Knowledge of Language (KoL) for any statistically-based Language Model (LM) must meet three requirements: (1) comprehensive coverage of linguistic phenomena; (2) replicable and stat…

Language ModellingSentence

Hard Examples Are All You Need: Maximizing GRPO Post-Training Under Annotation Budgets

2025-08-15 · Benjamin Pikus, Pratyush Ranjan Tiwari, Burton Ye arxiv

Collecting high-quality training examples for language model fine-tuning is expensive, with practical budgets limiting the amount of data that can be procured. We investigate whether example difficulty affects GRPO train…

Annotating picture description task responses for content analysis

2018-06-01 · WS 2018 6 · Levi King, Markus Dickinson

Given that all users of a language can be creative in their language usage, the overarching goal of this work is to investigate issues of variability and acceptability in written text, for both non-native speakers (NNSs)…

FormReading Comprehension