paper-with-me

Papers

Do We Need Neural Models to Explain Human Judgments of Acceptability?

2019-09-18 · Wang Jing, M. A. Kelly, David Reitter

Native speakers can judge whether a sentence is an acceptable instance of their language. Acceptability provides a means of evaluating whether computational language models are processing language in a human-like manner. We test the ability of computational language models, simple language features, and word embeddings to predict native English speakers judgments of acceptability on English-language essays written by non-native speakers. We find that much of the sentence acceptability variance can be captured by a combination of features including misspellings, word order, and word similarity (Pearson's r = 0.494). While predictive neural models fit acceptability judgments well (r = 0.527), we find that a 4-gram model with statistical smoothing is just as good (r = 0.528). Thanks to incorporating a count of misspellings, our 4-gram model surpasses both the previous unsupervised state-of-the art (Lau et al., 2015; r = 0.472), and the average non-expert native speaker (r = 0.46). Our results demonstrate that acceptability is well captured by n-gram statistics and simple language features.

📄 PDF Abstract BibTeX arXiv:1909.08663

Code (0)

등록된 구현이 없습니다.

Tasks

SentenceWord EmbeddingsWord Similarity

Similar Papers 제목 키워드 기반

Predicting Sentence Acceptability Judgments in Multimodal Contexts

2026-02-24 · Hyewon Jang, Nikolai Ilinykh, Sharid Loáiciga, Jey Han Lau 외 arxiv

Previous work has examined the capacity of deep neural networks (DNNs), particularly transformers, to predict human sentence acceptability judgments, both independently of context, and in document contexts. We consider t…

Reframing Human-AI Collaboration for Generating Free-Text Explanations

2021-12-16 · NAACL 2022 7 · Sarah Wiegreffe, Jack Hessel, Swabha Swayamdipta, Mark Riedl 외

Large language models are increasingly capable of generating fluent-appearing text with relatively little task-specific supervision. But can these models accurately explain classification decisions? We consider the task …

A Discerning Several Thousand Judgments: GPT-3 Rates the Article + Adjective + Numeral + Noun Construction

2023-01-29 · Kyle Mahowald

Knowledge of syntax includes knowledge of rare, idiosyncratic constructions. LLMs must overcome frequency biases in order to master such constructions. In this study, I prompt GPT-3 to give acceptability judgments on the…

CoLA

What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length

2024-11-04 · Lindia Tjuatja, Graham Neubig, Tal Linzen, Sophie Hao

When comparing the linguistic capabilities of language models (LMs) with humans using LM probabilities, factors such as the length of the sequence and the unigram frequency of lexical items have a significant effect on L…

Acceptable risks in Europe's proposed AI Act: Reasonableness and other principles for deciding how much risk management is enough

2023-07-26 · Henry Fraser, Jose-Miguel Bello y Villarino

This paper critically evaluates the European Commission's proposed AI Act's approach to risk management and risk acceptability for high-risk AI systems that pose risks to fundamental rights and safety. The Act aims to pr…

Management