paper-with-me

Papers

Digital Socrates: Evaluating LLMs through Explanation Critiques

2023-11-16 · Yuling Gu, Oyvind Tafjord, Peter Clark

While LLMs can provide reasoned explanations along with their answers, the nature and quality of those explanations are still poorly understood. In response, our goal is to define a detailed way of characterizing the explanation capabilities of modern models and to create a nuanced, interpretable explanation evaluation tool that can generate such characterizations automatically, without relying on expensive API calls or human annotations. Our approach is to (a) define the new task of explanation critiquing - identifying and categorizing any main flaw in an explanation and providing suggestions to address the flaw, (b) create a sizeable, human-verified dataset for this task, and (c) train an open-source, automatic critique model (called Digital Socrates) using this data. Through quantitative and qualitative analysis, we demonstrate how Digital Socrates is useful for revealing insights about student models by examining their reasoning chains, and how it can provide high-quality, nuanced, automatic evaluation of those model explanations for the first time. Digital Socrates thus fills an important gap in evaluation tools for understanding and improving the explanation behavior of models.

📄 PDF Abstract BibTeX arXiv:2311.09613

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations

2026-06-04 · Taewon Yun, Hyeonseong Park, Jeonghwan Choi, Hayoon Park 외 arxiv

Evaluating LLM mediators remains challenging, as mediation unfolds as a real-time trajectory shaped by disputants' shifting emotions, intentions, and context. Existing testbeds rely on a few expert-authored domains, vary…

SOCRATES: Simulation Optimization with Correlated Replicas and Adaptive Trajectory Evaluations

2025-11-01 · Haoting Zhang, Haoxian Chen, Donglin Zhan, Hanyang Zhao 외 arxiv

The field of simulation optimization (SO) encompasses various methods developed to optimize complex, expensive-to-sample stochastic systems. Established methods include, but are not limited to, ranking-and-selection for …

Epistemoverse: Toward an AI-Driven Knowledge Metaverse for Intellectual Heritage Preservation

2025-12-13 · Predrag K. Nikolić, Robert Prentner arxiv

Large language models (LLMs) have often been characterized as "stochastic parrots" that merely reproduce fragments of their training data. This study challenges that assumption by demonstrating that, when placed in an ap…

How Do LLMs Perform Two-Hop Reasoning in Context?

2025-02-19 · Tianyu Guo, Hanlin Zhu, Ruiqi Zhang, Jiantao Jiao 외

"Socrates is human. All humans are mortal. Therefore, Socrates is mortal." This classical example demonstrates two-hop reasoning, where a conclusion logically follows from two connected premises. While transformer-based …

Finetuning LLMs for Human Behavior Prediction in Social Science Experiments

2025-09-06 · Akaash Kolluri, Shengguang Wu, Joon Sung Park, Michael S. Bernstein arxiv

Large language models (LLMs) offer a powerful opportunity to simulate the results of social science experiments. In this work, we demonstrate that finetuning LLMs directly on individual-level responses from past experime…