Predicting the Difficulty of Language Proficiency Tests
Language proficiency tests are used to evaluate and compare the progress of language learners. We present an approach for automatic difficulty prediction of C-tests that performs on par with human experts. On the basis of detailed analysis of newly collected data, we develop a model for C-test difficulty introducing four dimensions: solution difficulty, candidate ambiguity, inter-gap dependency, and paragraph difficulty. We show that cues from all four dimensions contribute to C-test difficulty.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Textual complexity as a predictor of difficulty of listening items in language proficiency tests
In this paper we explore to what extent the difficulty of listening items in an English language proficiency test can be predicted by the textual properties of the prompt. We show that a system based on multiple text com…
Reading ComprehensionC-Test Collector: A Proficiency Testing Application to Collect Training Data for C-Tests
We present the C-Test Collector, a web-based tool that allows language learners to test their proficiency level using c-tests. Our tool collects anonymized data on test performance, which allows teachers to gain insights…
Jump-Starting Item Parameters for Adaptive Language Tests
A challenge in designing high-stakes language assessments is calibrating the test item difficulties, either a priori or from limited pilot test data. While prior work has addressed ‘cold start’ estimation of item difficu…
Language AcquisitionMulti-Task LearningSkills AssessmentFace Identification Proficiency Test Designed Using Item Response Theory
Measures of face-identification proficiency are essential to ensure accurate and consistent performance by professional forensic face examiners and others who perform face-identification tasks in applied scenarios. Curre…
Face IdentificationFace RecognitionMachine Learning--Driven Language Assessment
We describe a method for rapidly creating language proficiency assessments, and provide experimental evidence that such tests can be valid, reliable, and secure. Our approach is the first to use machine learning and natu…
BIG-bench Machine LearningLanguage AcquisitionSkills Assessmentvalid