paper-with-me

Papers

Semantic Geometry for policy-constrained interpretation

2025-12-10 · Nikit Phadke arxiv

We present a geometric framework for policy-constrained semantic interpretation that provably prevents hallucinated commitments in high-stakes domains. Semantic meaning is represented as direction on a unit sphere, evidence is modeled as sets of witness vectors, and admissible interpretations correspond to spherical convex regions. Policy constraints are introduced as explicit priors defined over the same manifold, separated from evidence geometry. Interpretation reduces to constrained optimization over admissible regions, with refusal emerging as a topologically necessary outcome under contradiction or policy exclusion. We connect this framework to information theory, Bayesian inference, and sheaf-theoretic semantics, proving that our complexity bounds are information-theoretically optimal. Empirical validation on large scale regulated financial data demonstrates zero hallucinated approvals across multiple policy regimes-the first such result at scale.

📄 PDF Abstract BibTeX arXiv:2512.14731

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian Inference

Similar Papers 제목 키워드 기반

Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space

2026-01-18 · Wang Zixian arxiv

We propose Orthogonalized Policy Optimization (OPO), a principled framework for large language model alignment derived from optimization in the Hilbert function space L2(pi_k). Lifting policy updates from the probability…

Embedding Safety into RL: A New Take on Trust Region Methods

2024-11-05 · Nikola Milosevic, Johannes Müller, Nico Scherf

Reinforcement Learning (RL) agents can solve diverse tasks but often exhibit unsafe behavior. Constrained Markov Decision Processes (CMDPs) address this by enforcing safety constraints, yet existing methods either sacrif…

Reinforcement Learning (RL)

Formal Semantic Geometry over Transformer-based Variational AutoEncoder

2022-10-12 · Yingji Zhang, Danilo S. Carvalho, Ian Pratt-Hartmann, André Freitas

Formal/symbolic semantics can provide canonical, rigid controllability and interpretability to sentence representations due to their \textit{localisation} or \textit{composition} property. How can we deliver such propert…

DisentanglementExplanation GenerationSentence

AlphaRoute: Large Language Models as Semantic Optimizers for Multi-Objective Routing

2026-07-22 · Kabir Murjani, Mishri Bhavsar, Manish I. Patel, Jonti Talukdar arxiv

Very Large Scale Integration (VLSI) global routing is an NP-hard combinatorial optimization problem requiring signal net assignment across capacity-constrained 3D grids while minimizing congestion, wirelength, and via tr…

Make it SING: Analyzing Semantic Invariants in Classifiers

2026-03-15 · Harel Yadid, Meir Yossef Levi, Roy Betser, Guy Gilboa arxiv

All classifiers, including state-of-the-art vision models, possess invariants, partially rooted in the geometry of their linear mappings. These invariants, which reside in the null-space of the classifier, induce equival…