paper-with-me

Papers

Is Solving Better Than Evaluating GenAI Solutions?

2026-07-30 · Ethan Dickey, Marios Mertzanidis, Alexandros Psomas arxiv

As Generative AI (GenAI) tools become increasingly capable of generating solutions to computing assignments, the computing education community is exploring pedagogical approaches that emphasize solution evaluation, verification, and critique alongside traditional solution generation. However, evidence regarding the impact of such evaluation-centered tasks on student learning remains limited, particularly in upper-division, theory-heavy courses. We conducted a randomized A/B crossover study (N=220) in a junior-level algorithms course to compare evaluating GenAI-generated solutions with traditional problem solving. Across six assignments, student working groups either solved challenging algorithmic problems directly or evaluated often-flawed GenAI-generated solutions, with roles reversed midway through the semester. We found no statistically significant differences between groups in midterm scores, final exam scores, overall course grades, or exam problems structurally aligned with the homework interventions. Students received significantly higher homework scores when evaluating GenAI-generated solutions, but this localized advantage did not translate into downstream summative gains. Survey data further indicated that most students reported no change in study habits in response to the intervention; however, those who reported adapting their study strategies rated the GenAI-evaluation assignments as significantly more helpful. These findings suggest that GenAI evaluation redistributes student effort from open-ended solution construction toward verification, diagnosis, and judgment, but does not automatically produce stronger conceptual transfer. We conclude that GenAI-evaluation activities can be incorporated into algorithms coursework without broad performance losses, but meaningful learning gains may require deliberate scaffolding that pushes students beyond simple error diagnosis.

📄 PDF Abstract BibTeX arXiv:2607.27586

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Widening Gap: The Benefits and Harms of Generative AI for Novice Programmers

2024-05-28 · James Prather, Brent Reeves, Juho Leinonen, Stephen MacNeil 외

Novice programmers often struggle through programming problem solving due to a lack of metacognitive awareness and strategies. Previous research has shown that novices can encounter multiple metacognitive difficulties wh…

Scaffold or Crutch? Examining College Students' Use and Views of Generative AI Tools for STEM Education

2024-12-03 · Karen D. Wang, Zhangyang Wu, L'Nard Tufts II, Carl Wieman 외

Developing problem-solving competency is central to Science, Technology, Engineering, and Mathematics (STEM) education, yet translating this priority into effective approaches to problem-solving instruction and assessmen…

Examining the Usage of Generative AI Models in Student Learning Activities for Software Programming

2025-11-17 · Rufeng Chen, Shuaishuai Jiang, Jiyun Shen, AJung Moon 외 arxiv

The rise of Generative AI (GenAI) tools like ChatGPT has created new opportunities and challenges for computing education. Existing research has primarily focused on GenAI's ability to complete educational tasks and its …

Explainable Generative AI (GenXAI): A Survey, Conceptualization, and Research Agenda

2024-04-15 · Johannes Schneider

Generative AI (GenAI) marked a shift from AI being able to recognize to AI being able to generate solutions for a wide variety of tasks. As the generated solutions and applications become increasingly more complex and mu…

Survey

Computational Hermeneutics: Evaluating generative AI as a cultural technology

2026-03-31 · Cody Kommers, Ruth Ahnert, Maria Antoniak, Emmanouil Benetos 외 arxiv

Generative AI systems are increasingly recognized as cultural technologies, yet current evaluation frameworks often treat culture as a variable to be measured rather than fundamental to the system's operation. Drawing on…