paper-with-me

홈 › Papers

Simulated Students in Tutoring Dialogues: Substance or Illusion?

2026-01-07 · Alexander Scarlatos, Jaewook Lee, Simon Woodhead, Andrew Lan arxiv

Advances in large language models (LLMs) enable many new innovations in education. However, evaluating the effectiveness of new technology requires real students, which is time-consuming and hard to scale up. Therefore, many recent works on LLM-powered tutoring solutions have used simulated students for both training and evaluation, often via simple prompting. Surprisingly, little work has been done to ensure or even measure the quality of simulated students. In this work, we formally define the student simulation task, propose a set of evaluation metrics that span linguistic, behavioral, and cognitive aspects, and benchmark a wide range of student simulation methods on these metrics. We experiment on a real-world math tutoring dialogue dataset, where both automated and human evaluation results show that prompting strategies for student simulation perform poorly; supervised fine-tuning and preference optimization yield much better but still limited performance, motivating future work on this challenging task.

📄 PDF Abstract BibTeX arXiv:2601.04025

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

How Real Is AI Tutoring? Comparing Simulated and Human Dialogues in One-on-One Instruction

2025-09-02 · Ruijia Li, Yuan-Hao Jiang, Jiatong Wang, Bo Jiang arxiv

Heuristic and scaffolded teacher-student dialogues are widely regarded as critical for fostering students' higher-order thinking and deep learning. However, large language models (LLMs) currently face challenges in gener…

Retrieval-Augmented Tutoring for Algorithm Tracing and Problem-Solving in AI Education

2026-05-13 · Mragisha Jain, Tirth Bhatt, Griffin Pitts, Aum Pandya 외 arxiv

Students learning algorithms often need support as they interpret traces, debug reasoning errors, and apply procedures across unfamiliar problem instances. In this paper, we present KITE (Knowledge-Informed Tutoring Engi…

Exploring LLMs for Predicting Tutor Strategy and Student Outcomes in Dialogues

2025-07-09 · Fareya Ikram, Alexander Scarlatos, Andrew Lan arxiv

Tutoring dialogues have gained significant attention in recent years, given the prominence of online learning and the emerging tutoring abilities of artificial intelligence (AI) agents powered by large language models (L…

Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors

2024-01-06 · Sanjit Kakarla, Danielle Thomas, Jionghao Lin, Shivang Gupta 외

Research suggests that tutors should adopt a strategic approach when addressing math errors made by low-efficacy students. Rather than drawing direct attention to the error, tutors should guide the students to identify a…

Math

Modeling Student Response Times: Towards Efficient One-on-one Tutoring Dialogues

2018-11-01 · WS 2018 11 · Luciana Benotti, Jayadev Bhaskaran, Sigtryggur Kjartansson, David Lang

In this paper we investigate the task of modeling how long it would take a student to respond to a tutor question during a tutoring dialogue. Solving such a task has applications in educational settings such as intellige…

Math