paper-with-me

홈 › Papers

Controlling Logical Collapse in LLMs via Algebraic Ontology Projection over F2

2026-05-13 · Hisashi Miyashita, Mgnite Inc arxiv

Do large language models internally encode ontological relations in a formally verifiable algebraic structure? We introduce Algebraic Ontology Projection (AOP), which projects LLM hidden states into the Galois Field F2 under Liskov Substitution Principle constraints, using only 42 relational pairs as algebraic keys. AOP achieves up to 93.33% zero-shot inclusion accuracy on unseen concept pairs (Gemma-2 Instruct with optimized prompt), with consistent 86.67% accuracy observed across multiple model families -- with no model tuning, but through prompt alone. This algebraic structure is strongly layer-dependent. We introduce Semantic Crystallisation (SC), a metric that quantifies F2 constraint satisfaction relative to a random baseline and predicts zero-shot accuracy without held-out data. System prompts act as algebraic boundary conditions: only their combination with instruction tuning prevents Late-layer Collapse -- a systematic degradation of logical consistency in the final layers, observed in 7 of 10 conditions. These findings reframe forward computation as an iterative process of algebraic organisation, and open a path toward LLMs whose logical structure is not merely approximated, but formally accessible.

📄 PDF Abstract BibTeX arXiv:2605.12968

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Geometric Price of Discrete Logic: Context-driven Manifold Dynamics of Number Representations

2026-03-24 · Long Zhang, Dai-jun Lin, Wei-neng Chen arxiv

Large language models (LLMs) generalize smoothly across continuous semantic spaces, yet strict logical reasoning demands the formation of discrete decision boundaries. Prevailing theories relying on linear isometric proj…

Logical Reasoning

Probing Ethical Framework Representations in Large Language Models: Structure, Entanglement, and Methodological Challenges

2026-03-24 · Weilun Xu, Alexander Rusnak, Frederic Kaplan arxiv

When large language models make ethical judgments, do their internal representations distinguish between normative frameworks, or collapse ethics into a single acceptability dimension? We probe hidden representations acr…

$\mathbfλ$-VAE: Variance Equalization for Posterior Collapse

2026-07-06 · Girum Demisse arxiv

Variational Autoencoders (VAEs) frequently suffer from posterior collapse, a failure mode in which the approximate posterior converges to the prior, rendering the latent code uninformative. Despite extensive research, a …

The Mercurial Top-Level Ontology of Large Language Models

2024-04-26 · Nele Köhler, Fabian Neuhaus

In our work, we systematize and analyze implicit ontological commitments in the responses generated by large language models (LLMs), focusing on ChatGPT 3.5 as a case study. We investigate how LLMs, despite having no exp…

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning

2026-04-04 · Zeyu Wang, Jingye Xu, Xiaogang Li, Peiyao Xiao 외 arxiv

Current multimodal benchmarks for scientific reasoning primarily evaluate local information extraction -- models recognize symbols and values and then perform textual inference. They do not assess whether models can reas…

Information ExtractionMultimodal Reasoning