paper-with-me

Papers

Lexical Bundle Frequency as a Construct-Relevant Candidate Feature in Automated Scoring of L2 Academic Writing

2025-04-11 · Burak Senel

Automated scoring (AS) systems are increasingly used for evaluating L2 writing, but require ongoing refinement for construct validity. While prior work suggested lexical bundles (LBs) - recurrent multi-word sequences satisfying certain frequency criteria - could inform assessment, their empirical integration into AS models needs further investigation. This study tested the impact of incorporating LB frequency features into an AS model for TOEFL independent writing tasks. Analyzing a sampled subcorpus (N=1,225 essays, 9 L1s) from the TOEFL11 corpus, scored by ETS-trained raters (Low, Medium, High), 3- to 9-word LBs were extracted, distinguishing prompt-specific from non-prompt types. A baseline Support Vector Machine (SVM) scoring model using established linguistic features (e.g., mechanics, cohesion, sophistication) was compared against an extended model including three aggregate LB frequency features (total prompt, total non-prompt, overall total). Results revealed significant, though generally small-effect, relationships between LB frequency (especially non-prompt bundles) and proficiency (p < .05). Mean frequencies suggested lower proficiency essays used more LBs overall. Critically, the LB-enhanced model improved agreement with human raters (Quadratic Cohen's Kappa +2.05%, overall Cohen's Kappa +5.63%), with notable gains for low (+10.1% exact agreement) and medium (+14.3% Cohen's Kappa) proficiency essays. These findings demonstrate that integrating aggregate LB frequency offers potential for developing more linguistically informed and accurate AS systems, particularly for differentiating developing L2 writers.

📄 PDF Abstract BibTeX arXiv:2504.08537

Code (1)

lbiap/main 공식 구현

Similar Papers 제목 키워드 기반

Phonological (un)certainty weights lexical activation

2017-11-17 · WS 2018 1 · Laura Gwilliams, David Poeppel, Alec Marantz, Tal Linzen

Spoken word recognition involves at least two basic computations. First is matching acoustic input to phonological categories (e.g. /b/, /p/, /d/). Second is activating words consistent with those phonological categories…

Lexical bundles in computational linguistics academic literature

2016-03-09 · Adel Rahimi

In this study we analyzed a corpus of 8 million words academic literature from Computational lingustics' academic literature. the lexical bundles from this corpus are categorized based on structures and functions.

LSBert: A Simple Framework for Lexical Simplification

2020-06-25 · Jipeng Qiang, Yun Li, Yi Zhu, Yunhao Yuan 외

Lexical simplification (LS) aims to replace complex words in a given sentence with their simpler alternatives of equivalent meaning, to simplify the sentence. Recently unsupervised lexical simplification approaches only …

Language ModelingLanguage ModellingLexical SimplificationSentence+1

Damping Properties of the Hair Bundle

2015-05-14

The viscous liquid surrounding a hair bundle dissipates energy and dampens oscillations, which poses a fundamental physical challenge to the high sensitivity and sharp frequency selectivity of hearing. To identify the me…

Heterotic String Model Building with Monad Bundles and Reinforcement Learning

2021-08-16 · Andrei Constantin, Thomas R. Harvey, Andre Lukas

We use reinforcement learning as a means of constructing string compactifications with prescribed properties. Specifically, we study heterotic SO(10) GUT models on Calabi-Yau three-folds with monad bundles, in search of …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)