paper-with-me

홈 › Papers

Predicting LLM Correctness in Prosthodontics Using Metadata and Hallucination Signals

2025-12-27 · Lucky Susanto, Anasta Pranawijayana, Cortino Sukotjo, Soni Prasad, Derry Wijaya arxiv

Large language models (LLMs) are increasingly adopted in high-stakes domains such as healthcare and medical education, where the risk of generating factually incorrect (i.e., hallucinated) information is a major concern. While significant efforts have been made to detect and mitigate such hallucinations, predicting whether an LLM's response is correct remains a critical yet underexplored problem. This study investigates the feasibility of predicting correctness by analyzing a general-purpose model (GPT-4o) and a reasoning-centric model (OSS-120B) on a multiple-choice prosthodontics exam. We utilize metadata and hallucination signals across three distinct prompting strategies to build a correctness predictor for each (model, prompting) pair. Our findings demonstrate that this metadata-based approach can improve accuracy by up to +7.14% and achieve a precision of 83.12% over a baseline that assumes all answers are correct. We further show that while actual hallucination is a strong indicator of incorrectness, metadata signals alone are not reliable predictors of hallucination. Finally, we reveal that prompting strategies, despite not affecting overall accuracy, significantly alter the models' internal behaviors and the predictive utility of their metadata. These results present a promising direction for developing reliability signals in LLMs but also highlight that the methods explored in this paper are not yet robust enough for critical, high-stakes deployment.

📄 PDF Abstract BibTeX arXiv:2512.22508

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Grad Detect: Gradient-Based Hallucination Detection in LLMs

2026-06-23 · Anand Kamat, Daniel Blake, Brent M. Werness arxiv

Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse tasks, yet they remain prone to generating hallucinations. Detecting these hallucinations is critical for deploying LLMs reliably in h…

Thinking, Faithful and Stable: Mitigating Hallucinations in LLMs

2025-11-19 · Chelsea Zou, Yiheng Yao, Basant Khalil arxiv

This project develops a self correcting framework for large language models (LLMs) that detects and mitigates hallucinations during multi-step reasoning. Rather than relying solely on final answer correctness, our approa…

Reinforcement Learning

Unsupervised Hallucination Detection by Inspecting Reasoning Processes

2025-09-12 · Ponhvoan Srey, Xiaobao Wu, Anh Tuan Luu arxiv

Unsupervised hallucination detection aims to identify hallucinated content generated by large language models (LLMs) without relying on labeled data. While unsupervised methods have gained popularity by eliminating labor…

How do Humans Process AI-generated Hallucination Contents: a Neuroimaging Study

2026-05-16 · Shuqi Zhu, Yi Zhong, Ziyi Ye, Bangde Du 외 arxiv

While AI-generated hallucinations pose considerable risks, the underlying cognitive mechanisms by which humans can successfully recognize or be misled by these hallucinations remain unclear. To address this problem, this…

Fact Verification

Were You Helpful -- Predicting Helpful Votes from Amazon Reviews

2024-12-03 · Emin Kirimlioglu, Harrison Kung, Dominic Orlando

This project investigates factors that influence the perceived helpfulness of Amazon product reviews through machine learning techniques. After extensive feature analysis and correlation testing, we identified key metada…

Sentiment Analysis