paper-with-me

홈 › Papers

Making LLMs Reliable When It Matters Most: A Five-Layer Architecture for High-Stakes Decisions

2025-11-10 · Alejandro R. Jadad arxiv

Current large language models (LLMs) excel in verifiable domains where outputs can be checked before action but prove less reliable for high-stakes strategic decisions with uncertain outcomes. This gap, driven by mutually reinforcing cognitive biases in both humans and artificial intelligence (AI) systems, threatens the defensibility of valuations and sustainability of investments in the sector. This report describes a framework emerging from systematic qualitative assessment across 7 frontier-grade LLMs and 3 market-facing venture vignettes under time pressure. Detailed prompting specifying decision partnership and explicitly instructing avoidance of sycophancy, confabulation, solution drift, and nihilism achieved initial partnership state but failed to maintain it under operational pressure. Sustaining protective partnership state required an emergent 7-stage calibration sequence, built upon a 4-stage initialization process, within a 5-layer protection architecture enabling bias self-monitoring, human-AI adversarial challenge, partnership state verification, performance degradation detection, and stakeholder protection. Three discoveries resulted: partnership state is achievable through ordered calibration but requires emergent maintenance protocols; reliability degrades when architectural drift and context exhaustion align; and dissolution discipline prevents costly pursuit of fundamentally wrong directions. Cross-model validation revealed systematic performance differences across LLM architectures. This approach demonstrates that human-AI teams can achieve cognitive partnership capable of preventing avoidable regret in high-stakes decisions, addressing return-on-investment expectations that depend on AI systems supporting consequential decision-making without introducing preventable cognitive traps when verification arrives too late.

📄 PDF Abstract BibTeX arXiv:2511.07669

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When Punctuation Matters: A Large-Scale Comparison of Prompt Robustness Methods for LLMs

2025-08-15 · Mikhail Seleznyov, Mikhail Chaichuk, Gleb Ershov, Alexander Panchenko 외 arxiv

Large Language Models (LLMs) are highly sensitive to subtle, non-semantic variations in prompt phrasing and formatting. In this work, we present the first systematic evaluation of 5 methods for improving prompt robustnes…

The Missing Knowledge Layer in AI: A Framework for Stable Human-AI Reasoning

2026-04-16 · Rikard Rosenbacke, Carl Rosenbacke, Victor Rosenbacke, Martin McKee arxiv

Large language models are increasingly integrated into decision-making in areas such as healthcare, law, finance, engineering, and government. Yet they share a critical limitation: they produce fluent outputs even when t…

InfoGatherer: Principled Information Seeking via Evidence Retrieval and Strategic Questioning

2026-03-06 · Maksym Taranukhin, Shuyue Stella Li, Evangelos Milios, Geoff Pleiss 외 arxiv

LLMs are increasingly deployed in high-stakes domains such as medical triage and legal assistance, often as document-grounded QA systems in which a user provides a description, relevant sources are retrieved, and an LLM …

AI Knows What's Wrong But Cannot Fix It: Helicoid Dynamics in Frontier LLMs Under High-Stakes Decisions

2026-03-12 · Alejandro R Jadad arxiv

Large language models perform reliably when their outputs can be checked: solving equations, writing code, retrieving facts. They perform differently when checking is impossible, as when a clinician chooses an irreversib…

Do LLMs Share Human-Like Biases? Causal Reasoning Under Prior Knowledge, Irrelevant Context, and Varying Compute Budgets

2026-02-03 · Hanna M. Dettki, Charley M. Wu, Bob Rehder arxiv

Large language models (LLMs) are increasingly used in domains where causal reasoning matters, yet it remains unclear whether their judgments reflect normative causal computation, human-like shortcuts, or brittle pattern …