paper-with-me

홈 › Papers

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

2026-05-14 · José Manuel de la Chica Rodríguez, Carlos Martí-González arxiv

Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal--agent failure: outputs can appear compliant without being compliant. Existing evaluation measures task accuracy but not whether governance constrains behaviour at the decision rationale level -- where regulated decisions must be auditable. We introduce five governance metrics that quantify policy compliance at the rationale level and apply them in a synthetic banking domain to compare text-only governance against mechanical enforcement: four primitives operating outside the model's interpretive loop. Under text-only governance, 27% of deferrals carry no decision-relevant information. Mechanical enforcement reduces this rate by 73%, more than doubles deferral information content, and raises task accuracy from MCC~$0.43$ to $0.88$. The improvement is driven by architectural separation: LLM-generated rationales under mechanical enforcement show comparable CDL to text-only governance -- the gain comes from removing clear-cut decisions from the model's control. A causal ablation confirms that each primitive is individually necessary. Our central finding is a governance-task decoupling: under structural stress, text-only governance degrades on both dimensions simultaneously, whereas mechanical enforcement preserves governance quality even as task performance drops. This implies that governance and task evaluation are distinct axes: accuracy is not a sufficient proxy for governance in regulated AI systems.

📄 PDF Abstract BibTeX arXiv:2605.14744

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A governance horizon for ethical-use constraints in open-weight AI models

2026-05-23 · Weiwei Xu, Hengzhi Ye, Haoran Ye, Kai Gao 외 arxiv

Ethical constraints on open-weight AI models are both a reflection of societal concerns and a foundation for AI governance policy. They are expected to propagate to downstream derivatives while implemented as voluntary m…

Governance-Aware Agent Telemetry for Closed-Loop Enforcement in Multi-Agent AI Systems

2026-04-06 · Anshul Pathak, Nishant Jain arxiv

Enterprise multi-agent AI systems produce thousands of inter-agent interactions per hour, yet existing observability tools capture these dependencies without enforcing anything. OpenTelemetry and Langfuse collect telemet…

Beyond Training: A Feasibility Taxonomy for Inference-Time AI Governance

2026-09-09 · Samar Ansari arxiv

Compute governance today is a governance of training: the thresholds, reporting requirements, and frontier-AI regimes now in force attach to training compute and treat the trained model as the regulatory unit. That pictu…

Adversarial Robustness

AGL-1: The Enterprise AI Governance Layer as a Control Plane for Trusted Enterprise Intelligence

2026-07-03 · Roopam W. Sure arxiv

Enterprise artificial intelligence is moving from isolated experimentation toward operational dependency across copilots, retrieval-augmented generation systems, autonomous agents, and AI-enabled business workflows. As t…

Ethical Hyper-Velocity (EHV): A Hardware-Rooted Zero-Trust Runtime Enforcement Architecture for Agentic AI Systems

2026-05-18 · Riddhi Mohan Sharma arxiv

As autonomous agentic systems scale across regulated critical infrastructures, the lack of mechanistic, hardware-rooted enforcement for high-frequency policy updates presents a fundamental safety gap. We present Ethical …