paper-with-me

Papers

Creating Large Language Model Resistant Exams: Guidelines and Strategies

2023-04-18 · Simon kaare Larsen

The proliferation of Large Language Models (LLMs), such as ChatGPT, has raised concerns about their potential impact on academic integrity, prompting the need for LLM-resistant exam designs. This article investigates the performance of LLMs on exams and their implications for assessment, focusing on ChatGPT's abilities and limitations. We propose guidelines for creating LLM-resistant exams, including content moderation, deliberate inaccuracies, real-world scenarios beyond the model's knowledge base, effective distractor options, evaluating soft skills, and incorporating non-textual information. The article also highlights the significance of adapting assessments to modern tools and promoting essential skills development in students. By adopting these strategies, educators can maintain academic integrity while ensuring that assessments accurately reflect contemporary professional settings and address the challenges and opportunities posed by artificial intelligence in education.

📄 PDF Abstract BibTeX arXiv:2304.12203

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation

2026-04-15 · Francesco Andrea Causio, Vittorio De Vita, Olivia Riccomi, Michele Ferramola 외 arxiv

While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when faced with non-English languages and multimodal diagnostic tasks. This …

Cross-Lingual TransferVisual Reasoning

Automatic Legal Writing Evaluation of LLMs

2025-04-29 · Ramon Pires, Roseval Malaquias Junior, Rodrigo Nogueira

Despite the recent advances in Large Language Models, benchmarks for evaluating legal writing remain scarce due to the inherent complexity of assessing open-ended responses in this domain. One of the key challenges in ev…

GPT-4 as an Agronomist Assistant? Answering Agriculture Exams Using Large Language Models

2023-10-10 · Bruno Silva, Leonardo Nunes, Roberto Estevão, Vijay Aski 외

Large language models (LLMs) have demonstrated remarkable capabilities in natural language understanding across various domains, including healthcare and finance. For some tasks, LLMs achieve similar or better performanc…

Information RetrievalManagementNatural Language UnderstandingRAG+2

Prompting Large Language Models for Supporting the Differential Diagnosis of Anemia

2024-09-20 · Elisa Castagnari, Lillian Muyama, Adrien Coulet

In practice, clinicians achieve a diagnosis by following a sequence of steps, such as laboratory exams, observations, or imaging. The pathways to reach diagnosis decisions are documented by guidelines authored by expert …

Decision MakingDiagnosticLanguage ModelingLanguage Modelling+1

SafeLLM: Extraction as a Hallucination-Resistant Alternative to Rewriting in Safety-Critical Settings

2026-06-11 · Julia Ive, Felix Jozsa, Evridiki Georgaki, Nabeel Sheikh 외 arxiv

Large language models (LLMs) are increasingly used to access organisational documentation, including standard operating procedures (SOPs), HR policies and institutional guidelines. However, retrieval-augmented generation…