paper-with-me

홈 › Papers

Interpretability from the Ground Up: Stakeholder-Centric Design of Automated Scoring in Educational Assessments

2025-11-21 · Yunsung Kim, Mike Hardy, Joseph Tey, Candace Thille, Chris Piech arxiv

AI-driven automated scoring systems offer scalable and efficient means of evaluating complex student-generated responses. Yet, despite increasing demand for transparency and interpretability, the field has yet to develop a widely accepted solution for interpretable automated scoring to be used in large-scale real-world assessments. This work takes a principled approach to address this challenge. We analyze the needs and potential benefits of interpretable automated scoring for various assessment stakeholder groups and develop four principles of interpretability -- (F)aithfulness, (G)roundedness, (T)raceability, and (I)nterchangeability (FGTI) -- targeted at those needs. To illustrate the feasibility of implementing these principles, we develop the AnalyticScore framework as a reference framework. When applied to the domain of text-based constructed-response scoring, AnalyticScore outperforms many uninterpretable scoring methods in terms of scoring accuracy and is, on average, within 0.06 QWK of the uninterpretable SOTA across 10 items from the ASAP-SAS dataset. By comparing against human annotators conducting the same featurization task, we further demonstrate that the featurization behavior of AnalyticScore aligns well with that of humans.

📄 PDF Abstract BibTeX arXiv:2511.17069

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Beyond Expertise and Roles: A Framework to Characterize the Stakeholders of Interpretable Machine Learning and their Needs

2021-01-24 · Harini Suresh, Steven R. Gomez, Kevin K. Nam, Arvind Satyanarayan

To ensure accountability and mitigate harm, it is critical that diverse stakeholders can interrogate black-box automated systems and find information that is understandable, relevant, and useful to them. In this paper, w…

DescriptiveInterpretable Machine Learning

SEED-SET: Scalable Evolving Experimental Design for System-level Ethical Testing

2026-03-02 · Anjali Parashar, Yingke Li, Eric Yang Yu, Fei Chen 외 arxiv

As autonomous systems such as drones, become increasingly deployed in high-stakes, human-centric domains, it is critical to evaluate the ethical alignment since failure to do so imposes imminent danger to human lives, an…

Gaussian Processes

Privacy Ethics Alignment in AI: A Stakeholder-Centric Based Framework for Ethical AI

2025-03-15 · Ankur Barthwal, Molly Campbell, Ajay Kumar Shrestha

The increasing integration of Artificial Intelligence (AI) in digital ecosystems has reshaped privacy dynamics, particularly for young digital citizens navigating data-driven environments. This study explores evolving pr…

Ethics

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

2026-08-02 · Muhammad Yousaf Rehman, Muhammad Islam arxiv

The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation, and risks to vulnerable online communities. Existing transformer-b…

Text Detection

On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs

2024-07-27 · Nitay Calderon, Roi Reichart

Recent advancements in NLP systems, particularly with the introduction of LLMs, have led to widespread adoption of these systems by a broad spectrum of users across various domains, impacting decision-making, the job mar…