paper-with-me

Papers

AgentCAT: Simulating Computerized Adaptive Testing via Multi-Agent Large Language Models

2026-06-20 · Weiyuan Zhou, Haiping Ma, Xiaoshan Yu, Changqian Wang, Shangshang Yang, Xingyi Zhang arxiv

Computerized Adaptive Testing (CAT), as a key technology for personalized education, aims to accurately assess examinee proficiency by retrieving exercises dynamically matching current ability estimates. However, existing CAT research is constrained by limitations of static offline data and isolated component optimization. Restricted by partial labels in offline logs, researchers degrade the dynamic assessment process into static sequence prediction. Current research focuses on isolated perspectives, e.g., selection or diagnosis, neglecting the overall CAT interaction process. To address this, we propose AgentCAT, a Large Language Model-based multi-agent simulation system, to construct a high-fidelity benchmarking environment for dynamic testing. This framework comprises three modules: (1) The examinee agent with memory retrieval and Chain-of-Thought reasoning simulates responses based on cognitive profiles; (2) The selection agent uses coarse-to-fine bucketing and knowledge graph exploration to balance local difficulty and global coverage; (3) The supervisor uses dual-auditing and robust update to ensure convergence and validity. To validate the framework, we evaluated on two real-world datasets across three dimensions: macro-level ability convergence, micro-level interaction logic, and data sparsity resilience. Results show AgentCAT achieves effective ability estimation, and its selection strategy balances difficulty adaptation and instructional coherence, aligning with human pedagogical intuition.

📄 PDF Abstract BibTeX arXiv:2606.21832

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Balancing Test Accuracy and Security in Computerized Adaptive Testing

2023-05-18 · Wanyong Feng, Aritra Ghosh, Stephen Sireci, Andrew S. Lan

Computerized adaptive testing (CAT) is a form of personalized testing that accurately measures students' knowledge levels while reducing test length. Bilevel optimization-based CAT (BOBCAT) is a recent framework that lea…

Bilevel OptimizationQuestion Selection

Survey of Computerized Adaptive Testing: A Machine Learning Perspective

2024-03-31 · Qi Liu, Yan Zhuang, Haoyang Bi, Zhenya Huang 외

Computerized Adaptive Testing (CAT) provides an efficient and tailored method for assessing the proficiency of examinees, by dynamically adjusting test questions based on their performance. Widely adopted across diverse …

cognitive diagnosisQuestion SelectionSociologySurvey

BOBCAT: Bilevel Optimization-Based Computerized Adaptive Testing

2021-08-17 · Aritra Ghosh, Andrew Lan

Computerized adaptive testing (CAT) refers to a form of tests that are personalized to every student/test taker. CAT methods adaptively select the next most informative question/item for each student given their response…

Bilevel OptimizationQuestion Selection

AgentCAT: An LLM Agent for Extracting and Analyzing Catalytic Reaction Data from Chemical Engineering Literature

2026-02-10 · Wei Yang, Zihao Liu, Tao Tan, Xiao Hu 외 arxiv

This paper presents a large language model (LLM) agent named AgentCAT, which extracts and analyzes catalytic reaction data from chemical engineering papers, %and supports natural language based interactive analysis of th…

An Intelligent Testing Strategy for Vocabulary Assessment of Chinese Second Language Learners

2019-08-01 · WS 2019 8 · Wei Zhou, Renfen Hu, Feipeng Sun, Ronghuai Huang

Vocabulary is one of the most important parts of language competence. Testing of vocabulary knowledge is central to research on reading and language. However, it usually costs a large amount of time and human labor to bu…