paper-with-me

홈 › Papers

Mitigating Learnerese Effects for CEFR Classification

2022-07-01 · NAACL (BEA) 2022 7 · Rricha Jalota, Peter Bourgonje, Jan Van Sas, Huiyan Huang

The role of an author’s L1 in SLA can be challenging for automated CEFR classification, in that texts from different L1 groups may be too heterogeneous to combine them as training data. We experiment with recent debiasing approaches by attempting to devoid textual representations of L1 features. This results in a more homogeneous group when aggregating CEFR-annotated texts from different L1 groups, leading to better classification performance. Using iterative null-space projection, we marginally improve classification performance for a linear classifier by 1 point. An MLP (e.g. non-linear) classifier remains unaffected by this procedure. We discuss possible directions of future work to attempt to increase this performance gain.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Classification

Similar Papers 제목 키워드 기반

Orthogonal Series Estimation for the Ratio of Conditional Expectation Functions

2022-12-26 · Kazuhiko Shinoda, Takahiro Hoshino

In various fields of data science, researchers are often interested in estimating the ratio of conditional expectation functions (CEFR). Specifically in causal inference problems, it is sometimes natural to consider rati…

Causal Inference

An Effective Automated Speaking Assessment Approach to Mitigating Data Scarcity and Imbalanced Distribution

2024-04-11 · Tien-Hong Lo, Fu-An Chao, Tzu-I Wu, Yao-Ting Sung 외

Automated speaking assessment (ASA) typically involves automatic speech recognition (ASR) and hand-crafted feature extraction from the ASR transcript of a learner's speech. Recently, self-supervised learning (SSL) has sh…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Self-Supervised Learningspeech-recognition+1

Experiments with Universal CEFR Classification

2018-04-18 · WS 2018 6 · Sowmya Vajjala, Taraka Rama

The Common European Framework of Reference (CEFR) guidelines describe language proficiency of learners on a scale of 6 levels. While the description of CEFR guidelines is generic across languages, the development of auto…

ClassificationGeneral Classification

UniversalCEFR: Enabling Open Multilingual Research on Language Proficiency Assessment

2025-06-02 · Joseph Marvin Imperial, Abdullah Barayan, Regina Stodden, Rodrigo Wilkens 외

We introduce UniversalCEFR, a large-scale multilingual multidimensional dataset of texts annotated according to the CEFR (Common European Framework of Reference) scale in 13 languages. To enable open research in both aut…

Enhancing Marker Scoring Accuracy through Ordinal Confidence Modelling in Educational Assessments

2025-05-29 · Abhirup Chakravarty, Mark Brenchley, Trevor Breakspear, Ian Lewin 외

A key ethical challenge in Automated Essay Scoring (AES) is ensuring that scores are only released when they meet high reliability standards. Confidence modelling addresses this by assigning a reliability estimate measur…

Automated Essay Scoring