paper-with-me

홈 › Papers

When AI Fails, What Works? A Data-Driven Taxonomy of Real-World AI Risk Mitigation Strategies

2026-03-04 · Evgenija Popchanovska, Ana Gjorgjevikj, Maryan Rizinski, Lubomir Chitkushev, Irena Vodenska, Dimitar Trajanov arxiv

Large language models (LLMs) are increasingly embedded in high-stakes workflows, where failures propagate beyond isolated model errors into systemic breakdowns that can lead to legal exposure, reputational damage, and material financial losses. Building on this shift from model-centric risks to end-to-end system vulnerabilities, we analyze real-world AI incident reporting and mitigation actions to derive an empirically grounded taxonomy that links failure dynamics to actionable interventions. Using a unified corpus of 9,705 media-reported AI incident articles, we extract explicit mitigation actions from 6,893 texts via structured prompting and then systematically classify responses to extend MIT's AI Risk Mitigation Taxonomy. Our taxonomy introduces four new mitigation categories, including 1) Corrective and Restrictive Actions, 2) Legal/Regulatory and Enforcement Actions, 3) Financial, Economic, and Market Controls, and 4) Avoidance and Denial, capturing response patterns that are becoming increasingly prevalent as AI deployment and regulation evolve. Quantitatively, we label the mitigation dataset with 32 distinct labels, producing 23,994 label assignments; 9,629 of these reflect previously unseen mitigation patterns, yielding a 67% increase of the original subcategory coverage and substantially enhancing the taxonomy's applicability to emerging systemic failure modes. By structuring incident responses, the paper strengthens "diagnosis-to-prescription" guidance and advances continuous, taxonomy-aligned post-deployment monitoring to prevent cascading incidents and downstream impact.

📄 PDF Abstract BibTeX arXiv:2603.04259

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RECALL: Library-Like Behavior In Language Models is Enhanced by Self-Referencing Causal Cycles

2025-01-23 · Munachiso Nwadike, Zangir Iklassov, Toluwani Aremu, Tatsuya Hiraoka 외

We introduce the concept of the self-referencing causal cycle (abbreviated RECALL) - a mechanism that enables large language models (LLMs) to bypass the limitations of unidirectional causality, which underlies a phenomen…

When to Align, When to Predict: A Phase Diagram for Multimodal Learning

2026-06-09 · Ilay Kamai, Hugues Van Assel, Aviv Regev, Hagai B. Perets 외 arxiv

Cross-modal alignment (CA) and cross-modal prediction (CP) are the dominant paradigms for multimodal representation learning, yet there is no systematic understanding of when each succeeds, when each fails, and when cros…

Representation Learning

Programming with Data: Test-Driven Data Engineering for Self-Improving LLMs from Raw Corpora

2026-04-27 · Chenkai Pan, Xinglong Xu, Yuhang Xu, Yujun Wu 외 arxiv

Reliably transferring specialized human knowledge from text into large language models remains a fundamental challenge in artificial intelligence. Fine-tuning on domain corpora has enabled substantial capability gains, b…

A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis

2026-05-02 · Muhammad Arslan Manzoor, Dilshod Azizov, Daniil Orel, Umer Siddique 외 arxiv

News outlets shape public opinion at a scale that makes automated detection of political bias and factuality essential. However, the field still lacks unified resources, comprehensive evaluations across diverse approache…

Reinforcement Learning

What's Left Unsaid? Detecting and Correcting Misleading Omissions in Multimodal News Previews

2026-01-09 · Fanxiao Li, Jiaying Wu, Tingchao Fu, Dayang Li 외 arxiv

Even when factually correct, social-media news previews (image-headline pairs) can induce interpretation drift: by selectively omitting crucial context, they lead readers to form judgments that diverge from what the full…