paper-with-me

홈 › Papers

CARE: Privacy-Compliant Agentic Reasoning with Evidence Discordance

2026-04-01 · Haochen Liu, Weien Li, Rui Song, Zeyu Li, Chun Jason Xue, Xiao-Yang Liu, Sam Nallaperuma, Xue Liu, Ye Yuan arxiv

Large language model (LLM) systems are increasingly used to support high-stakes decision-making, but they typically perform worse when the available evidence is internally inconsistent. Such a scenario exists in real-world healthcare settings, with patient-reported symptoms contradicting medical signs. To study this problem, we introduce MIMIC-DOS, a dataset for short-horizon organ dysfunction worsening prediction in the intensive care unit (ICU) setting. We derive this dataset from the widely recognized MIMIC-IV, a publicly available electronic health record dataset, and construct it exclusively from cases in which discordance between signs and symptoms exists. This setting poses a substantial challenge for existing LLM-based approaches, with single-pass LLMs and agentic pipelines often struggling to reconcile such conflicting signals. To address this problem, we propose CARE: a multi-stage privacy-compliant agentic reasoning framework in which a proprietary LLM provides guidance by generating structured categories and transitions without accessing sensitive patient data, while a local LLM uses these categories and transitions to support evidence acquisition and final decision-making. Empirically, under controlled retrospective evaluation on MIMIC-DOS, CARE achieves the best overall performance across key metrics among the evaluated LLMs and agentic workflows, showing that it can more robustly handle conflicting clinical evidence while preserving privacy.

📄 PDF Abstract BibTeX arXiv:2604.01113

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards a HIPAA Compliant Agentic AI System in Healthcare

2025-04-24 · Subash Neupane, Sudip Mittal, Shahram Rahimi

Agentic AI systems powered by Large Language Models (LLMs) as their foundational reasoning engine, are transforming clinical workflows such as medical report generation and clinical summarization by autonomously analyzin…

AttributeMedical Report Generation

CARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning with an Evidence-Grounded Agentic Framework

2026-03-02 · Yuexi Du, Jinglu Wang, Shujie Liu, Nicha C. Dvornek 외 arxiv

Large visual language models (VLMs) have shown strong multi-modal medical reasoning ability, but most operate as end-to-end black boxes, diverging from clinicians' evidence-based, staged workflows and hindering clinical …

Reinforcement LearningVisual Grounding

A T-API-Compliant ReAct Agentic Loop for Optical Networks: Generic vs. Domain-Specific Tool Abstractions

2026-06-16 · Seyed Morteza Ahmadian, Paolo Monti, Carlos Natalino arxiv

Optical networks need intent-driven, closed-loop agentic management, a key enabler for higher autonomy levels. We present the first T-API-compliant reasoning and act (ReAct) loop. We show that domain-specific composite t…

Agentic-AI Healthcare: Multilingual, Privacy-First Framework with MCP Agents

2025-09-25 · Mohammed A. Shehab arxiv

This paper introduces Agentic-AI Healthcare, a privacy-aware, multilingual, and explainable research prototype developed as a single-investigator project. The system leverages the emerging Model Context Protocol (MCP) to…

FCMBench: The First Large-scale Financial Credit Multimodal Benchmark for Real-world Applications

2026-01-01 · Yehui Yang, Dalu Yang, Fangxin Shang, Wenshuo Zhou 외 arxiv

FCMBench is the first large-scale and privacy-compliant multimodal benchmark for real-world financial credit applications, covering tasks and robustness challenges from domain specific workflows and constraints. The curr…