paper-with-me

홈 › Papers

Taxonomy of Comprehensive Safety for Clinical Agents

2025-09-26 · Jean Seo, Hyunkyung Lee, Gibaeg Kim, Wooseok Han, Jaehyo Yoo, Seungseop Lim, Kihun Shin, Eunho Yang arxiv

Safety is a paramount concern in clinical chatbot applications, where inaccurate or harmful responses can lead to serious consequences. Existing methods--such as guardrails and tool calling--often fall short in addressing the nuanced demands of the clinical domain. In this paper, we introduce TACOS (TAxonomy of COmprehensive Safety for Clinical Agents), a fine-grained, 21-class taxonomy that integrates safety filtering and tool selection into a single user intent classification step. TACOS is a taxonomy that can cover a wide spectrum of clinical and non-clinical queries, explicitly modeling varying safety thresholds and external tool dependencies. To validate our taxonomy, we curate a TACOS-annotated dataset and perform extensive experiments. Our results demonstrate the value of a new taxonomy specialized for clinical agent settings, and reveal useful insights about train data distribution and pretrained knowledge of base models.

📄 PDF Abstract BibTeX arXiv:2509.22041

Code (0)

등록된 구현이 없습니다.

Tasks

Intent Classification

Similar Papers 제목 키워드 기반

AgentOps: Enabling Observability of LLM Agents

2024-11-08 · Liming Dong, Qinghua Lu, Liming Zhu

Large language model (LLM) agents have demonstrated remarkable capabilities across various domains, gaining extensive attention from academia and industry. However, these agents raise significant concerns on AI safety du…

AI AgentLanguage ModelingLanguage ModellingLarge Language Model+1

Swiss Cheese Model for AI Safety: A Taxonomy and Reference Architecture for Multi-Layered Guardrails of Foundation Model Based Agents

2024-08-05 · Md Shamsujjoha, Qinghua Lu, Dehai Zhao, Liming Zhu

Foundation Model (FM)-based agents are revolutionizing application development across various domains. However, their rapidly growing capabilities and autonomy have raised significant concerns about AI safety. Researcher…

modelSystematic Literature Review

MATRIX: Multi-Agent simulaTion fRamework for safe Interactions and conteXtual clinical conversational evaluation

2025-08-26 · Ernest Lim, Yajie Vera He, Jared Joselowitz, Kate Preston 외 arxiv

Despite the growing use of large language models (LLMs) in clinical dialogue systems, existing evaluations focus on task completion or fluency, offering little insight into the behavioral and risk management requirements…

A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?

2025-05-16 · Ada Chen, Yongjiang Wu, Junyuan Zhang, Jingyu Xiao 외

Recently, AI-driven interactions with computing devices have advanced from basic prototype tools to sophisticated, LLM-based systems that emulate human-like operations in graphical user interfaces. We are now witnessing …

MPIB: A Benchmark for Medical Prompt Injection Attacks and Clinical Safety in LLMs

2026-02-06 · Junhyeok Lee, Han Jang, Kyu Sung Choi arxiv

Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) systems are increasingly integrated into clinical workflows; however, prompt injection attacks can steer these systems toward clinically unsafe or mis…