paper-with-me

홈 › Papers

CareTransition-Audit: A Benchmark to Audit Discharge Summaries for Efficient Care Transitions

2026-04-07 · Akshat Dasula, Prasanna Desikan, Jaideep Srivastava, Shivali Dalmia, Abhishek Mukherji arxiv

Incomplete or inconsistent discharge documentation drives care fragmentation and avoidable readmissions. Despite its critical role in patient safety, auditing discharge summaries relies on manual review and does not scale. We propose an automated framework for auditing discharge summaries using large language models (LLMs). Our approach operationalizes the DISCHARGED framework into a checklist of 46 questions. Using 50 summaries from the MIMIC-IV database, with clinician ground-truth labels, we benchmark 11 LLMs. Model-assessed mean documentation completeness ranges from 54.9% to 74.2%, and the best-performing models achieve a Cohen's kappa values around 0.5 against clinician labels, indicating moderate agreement. All models struggle to identify ambiguous documentation (Unclear), highlighting a key gap in current automated auditing. This work provides a clinician-validated benchmark and zero-shot baselines for systematic quality improvement in clinical documentation.

📄 PDF Abstract BibTeX arXiv:2604.05435

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Planner-Auditor Twin: Agentic Discharge Planning with FHIR-Based LLM Planning, Guideline Recall, Optional Caching and Self-Improvement

2026-01-28 · Kaiyuan Wu, Aditya Nagori, Rishikesan Kamaleswaran arxiv

Objective: Large language models (LLMs) show promise for clinical discharge planning, but their use is constrained by hallucination, omissions, and miscalibrated confidence. We introduce a self-improving, cache-optional …

VERIRAG: A Post-Retrieval Auditing of Scientific Study Summaries

2025-07-23 · Shubham Mohole, Hongjun Choi, Shusen Liu, Christine Klymko 외 arxiv

Can democratized information gatekeepers and community note writers effectively decide what scientific information to amplify? Lacking domain expertise, such gatekeepers rely on automated reasoning agents that use RAG to…

Vulnerability Detection

Eidos: An Open-Source Auditory Periphery Modeling Toolkit and Evaluation of Cross-Lingual Phonemic Contrasts

2020-05-01 · LREC 2020 5 · Alex Gutkin, er

Many analytical models that mimic, in varying degree of detail, the basic auditory processes involved in human hearing have been developed over the past decades. While the auditory periphery mechanisms responsible for tr…

CW-B: Class Weighted Boosting Framework for Imbalance Resilient Multi Class Cardiac Phenotyping

2026-06-29 · Sijia Li, Xiaoyu Tan, Chen Zhan, Yuanji Ma 외 arxiv

Cardiac discharge phenotyping informs post-discharge treatment and follow-up, but real-world records are often incomplete and class-imbalanced, increasing the risk of missed high-risk phenotypes. We propose CW-B, a clini…

A Computational Audit of Demographic Association Encoding in ClinicalBERT Language Predictions

2026-06-12 · Kehinde Temitayo Soetan arxiv

Transformer-based clinical language models are increasingly integrated into high-stakes clinical decision support pipelines, yet the computational mechanisms through which demographic associations encoded in medical docu…