paper-with-me

Linguistic Acceptability

5개 벤치마크 · 논문 82편 · 이 태스크의 논문 보기 →

Benchmarks

CoLA

결과 43개

RuCoLA

결과 9개

CoLA Dev

결과 6개

ItaCoLA

결과 4개

DaLAJ

결과 1개

Most implemented

Big Bird: Transformers for Longer Sequences

2020-07-28 · 구현 14개

Papers

Data filtering methods for training language models

2026-05-28 · Egor Shevchenko, Elena Bruches arxiv

Data quality is a critical factor in the effectiveness of machine learning models. Label errors, present even in widely used benchmarks, introduce noise into training data and reduce model generalization. In this work, w…

Linguistic AcceptabilityEmotion ClassificationLabel Error DetectionText Classification

DunbaaBERT: From Sacrifice to Semantics

2026-05-26 · Iffat Maab, Waleed Jamil, Raphael Schmitt arxiv

Large language models have achieved strong performance across many NLP tasks, yet Urdu remains comparatively underexplored due to limited resources and fragmented evaluation settings. To address this gap, we introduce Du…

Linguistic AcceptabilityNews ClassificationSentiment Analysis

Dialects of Translationese Shape Language Model Learning

2026-02-18 · Jenny Kunz arxiv

Machine-translated data is widely used in multilingual NLP, particularly where native text is scarce. However, translated text differs systematically from native text. This phenomenon is known as translationese, and it r…

Linguistic AcceptabilityLanguage Modelling

Preferences for Idiomatic Language are Acquired Slowly -- and Forgotten Quickly: A Case Study on Swedish

2026-02-03 · Jenny Kunz arxiv

In this study, we investigate how language models develop preferences for \textit{idiomatic} as compared to \textit{linguistically acceptable} Swedish, both during pretraining and when adapting a model from English to Sw…

Linguistic Acceptability

DaLA: Danish Linguistic Acceptability Evaluation Guided by Real World Errors

2025-12-04 · Gianluca Barmina, Nathalie Carmen Hau Norman, Peter Schneider-Kamp, Lukas Galke Poech arxiv

We present an enhanced benchmark for evaluating linguistic acceptability in Danish. We first analyze the most common errors found in written Danish. Based on this analysis, we introduce a set of fourteen corruption funct…

Linguistic Acceptability

SindBERT, the Sailor: Charting the Seas of Turkish NLP

2025-10-24 · Raphael Schmitt, Stefan Schweter arxiv

Transformer models have revolutionized NLP, yet many morphologically rich languages remain underrepresented in large-scale pre-training efforts. With SindBERT, we set out to chart the seas of Turkish NLP, providing the f…

Linguistic AcceptabilityPart-Of-Speech Tagging

전체 82편 보기 →