paper-with-me

홈 › Papers

ProQE: Proficiency-wise Quality Estimation dataset for Grammatical Error Correction

2022-06-01 · LREC 2022 6 · Yujin Takahashi, Masahiro Kaneko, Masato Mita, Mamoru Komachi

This study investigates how supervised quality estimation (QE) models of grammatical error correction (GEC) are affected by the learners’ proficiency with the data. QE models for GEC evaluations in prior work have obtained a high correlation with manual evaluations. However, when functioning in a real-world context, the data used for the reported results have limitations because prior works were biased toward data by learners with relatively high proficiency levels. To address this issue, we created a QE dataset that includes multiple proficiency levels and explored the necessity of performing proficiency-wise evaluation for QE of GEC. Our experiments demonstrated that differences in evaluation dataset proficiency affect the performance of QE models, and proficiency-wise evaluation helps create more robust models.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Grammatical Error Correction

Similar Papers 제목 키워드 기반

Proficiency Matters Quality Estimation in Grammatical Error Correction

2022-01-17 · Yujin Takahashi, Masahiro Kaneko, Masato Mita, Mamoru Komachi

This study investigates how supervised quality estimation (QE) models of grammatical error correction (GEC) are affected by the learners' proficiency with the data. QE models for GEC evaluations in prior work have obtain…

Grammatical Error Correction

Progressive Query Expansion for Retrieval Over Cost-constrained Data Sources

2024-06-11 · Muhammad Shihab Rashid, Jannat Ara Meem, Yue Dong, Vagelis Hristidis

Query expansion has been employed for a long time to improve the accuracy of query retrievers. Earlier works relied on pseudo-relevance feedback (PRF) techniques, which augment a query with terms extracted from documents…

HallucinationRetrieval

PsyScore: A Psychometrically-Aware Framework for Trait-Adaptive Essay Scoring and ZPD-Scaffolded Feedback

2026-06-18 · Wei Xia, Jin Wu, Haoran Shi, Xiangyu Wang 외 arxiv

Effective Automated Essay Scoring (AES) are expected to support both reliable assessment and actionable instructional feedback. However, existing approaches often treat scoring and feedback as separate components: neural…

Automated Essay Scoring

MENLO: From Preferences to Proficiency -- Evaluating and Modeling Native-like Quality Across 47 Languages

2025-09-30 · Chenxi Whitehouse, Sebastian Ruder, Tony Lin, Oksana Kurylo 외 arxiv

Ensuring native-like quality of large language model (LLM) responses across many languages is challenging. To address this, we introduce MENLO, a framework that operationalizes the evaluation of native-like response qual…

Reinforcement LearningMulti-Task Learning

CuriosAI Submission to the EgoExo4D Proficiency Estimation Challenge 2025

2025-07-08 · Hayato Tanoue, Hiroki Nishihara, Yuma Suzuki, Takayuki Hori 외 arxiv

This report presents the CuriosAI team's submission to the EgoExo4D Proficiency Estimation Challenge at CVPR 2025. We propose two methods for multi-view skill assessment: (1) a multi-task learning framework using Sapiens…

Multi-Task Learning