paper-with-me

Papers

How critically can an AI think? A framework for evaluating the quality of thinking of generative artificial intelligence

2024-06-20 · Luke Zaphir, Jason M. Lodge, Jacinta Lisec, Dom McGrath, Hassan Khosravi

Generative AI such as those with large language models have created opportunities for innovative assessment design practices. Due to recent technological developments, there is a need to know the limits and capabilities of generative AI in terms of simulating cognitive skills. Assessing student critical thinking skills has been a feature of assessment for time immemorial, but the demands of digital assessment create unique challenges for equity, academic integrity and assessment authorship. Educators need a framework for determining their assessments vulnerability to generative AI to inform assessment design practices. This paper presents a framework that explores the capabilities of the LLM ChatGPT4 application, which is the current industry benchmark. This paper presents the Mapping of questions, AI vulnerability testing, Grading, Evaluation (MAGE) framework to methodically critique their assessments within their own disciplinary contexts. This critique will provide specific and targeted indications of their questions vulnerabilities in terms of the critical thinking skills. This can go on to form the basis of assessment design for their tasks.

📄 PDF Abstract BibTeX arXiv:2406.14769

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Argument Reconstruction as Supervision for Critical Thinking in LLMs

2026-03-18 · Hyun Ryu, Gyouk Chu, Gregor Betz, Eunho Yang 외 arxiv

To think critically about arguments, human learners are trained to identify, reconstruct, and evaluate arguments. Argument reconstruction is especially important because it makes an argument's underlying inferences expli…

Test-Time Deep Thinking to Explore Implicit Rules

2026-05-24 · Wentong Chen, Xin Cong, Zhong Zhang, Yaxi Lu 외 arxiv

With the continuous advancement of Large Language Models (LLMs), intelligent agents are becoming increasingly vital. However, these agents often fail in environments governed by implicit rules--hidden constraints that ca…

Reinforcement Learning

Cognitive Profiling of LRMs' Reasoning Traces Using Bloom's Taxonomy

2026-08-24 · Maria-Eleni Zoumpoulidi, Georgios Paraskevopoulos, Alexandros Potamianos arxiv

Large Reasoning Models (LRMs) have revolutionized reasoning in LLMs, and the increasing public availability of reasoning traces creates valuable opportunities to study model behavior not only at the surface level but als…

THINK-Bench: Evaluating Thinking Efficiency and Chain-of-Thought Quality of Large Reasoning Models

2025-05-28 · Zhiyuan Li, Yi Chang, Yuan Wu

Large reasoning models (LRMs) have achieved impressive performance in complex tasks, often outperforming conventional large language models (LLMs). However, the prevalent issue of overthinking severely limits their compu…

Computational Efficiency

Pilot Study on Generative AI and Critical Thinking in Higher Education Classrooms

2025-08-29 · W. F. Lamberti, S. R. Lawrence, D. White, S. Kim 외 arxiv

Generative AI (GAI) tools have seen rapid adoption in educational settings, yet their role in fostering critical thinking remains underexplored. While previous studies have examined GAI as a tutor for specific lessons or…