paper-with-me

홈 › Papers

Are You Doubtful? Oh, It Might Be Difficult Then! Exploring the Use of Model Uncertainty for Question Difficulty Estimation

2024-12-16 · Leonidas Zotos, Hedderik van Rijn, Malvina Nissim

In an educational setting, an estimate of the difficulty of multiple-choice questions (MCQs), a commonly used strategy to assess learning progress, constitutes very useful information for both teachers and students. Since human assessment is costly from multiple points of view, automatic approaches to MCQ item difficulty estimation are investigated, yielding however mixed success until now. Our approach to this problem takes a different angle from previous work: asking various Large Language Models to tackle the questions included in three different MCQ datasets, we leverage model uncertainty to estimate item difficulty. By using both model uncertainty features as well as textual features in a Random Forest regressor, we show that uncertainty features contribute substantially to difficulty prediction, where difficulty is inversely proportional to the number of students who can correctly answer a question. In addition to showing the value of our approach, we also observe that our model achieves state-of-the-art results on the USMLE and CMCQRD publicly available datasets.

📄 PDF Abstract BibTeX arXiv:2412.11831

Code (0)

등록된 구현이 없습니다.

Tasks

Multiple-choice

Similar Papers 제목 키워드 기반

Can Model Uncertainty Function as a Proxy for Multiple-Choice Question Item Difficulty?

2024-07-07 · Leonidas Zotos, Hedderik van Rijn, Malvina Nissim

Estimating the difficulty of multiple-choice questions would be great help for educators who must spend substantial time creating and piloting stimuli for their tests, and for learners who want to practice. Supervised ap…

Multiple-choice

Tribrid: Stance Classification with Neural Inconsistency Detection

2021-09-14 · EMNLP 2021 11 · Song Yang, Jacopo Urbani

We study the problem of performing automatic stance classification on social media with neural architectures such as BERT. Although these architectures deliver impressive results, their level is not yet comparable to the…

ClassificationFact CheckingStance Classification

Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents

2026-02-18 · Wenxuan Ding, Nicholas Tomlin, Greg Durrett arxiv

LLM agents are deployed in environments where they must interact to acquire information. In these scenarios, the agent must reason about inherent cost-uncertainty tradeoffs in how to act, such as when to stop exploring a…

Development of a classifiers/quantifiers dictionary towards French-Japanese MT

2019-02-21 · MTSummit 2017 9 · Mutsuko Tomokiyo, Mathieu Mangeot, Christian Boitet

Although classifiers/quantifiers (CQs) expressions appear frequently in everyday communications or written documents, they are described neither in classical bilingual paper dictionaries , nor in machine-readable diction…

Machine TranslationTranslation

Exploring the Limits of Epistemic Uncertainty Quantification in Low-Shot Settings

2021-11-18 · NeurIPS Workshop LatinX_in_AI 2021 12 · Matias Valdenegro-Toro

Uncertainty quantification in neural network promises to increase safety of AI systems, but it is not clear how performance might vary with the training set size. In this paper we evaluate seven uncertainty methods on Fa…

Out-of-Distribution DetectionUncertainty Quantification