paper-with-me

홈 › Papers

Internalizing Tools as Morphisms in Graded Transformers

2025-11-21 · Tony Shaska arxiv

We introduce a graded formulation of internal symbolic computation for transformers. The hidden space is endowed with a grading $V=\bigoplus_{g\in G}V_g$, and symbolic operations are realized as typed block maps (morphisms) $φ_{h\leftarrow g}:V_g\to V_h$ that are activated selectively by a differentiable routing policy. A self-supervised \emph{graded utility functional}, defined as the loss reduction induced by a candidate morphism, governs activation and yields sparse, interpretable behavior. We develop the algebraic and geometric foundations: an internal model category whose objects are homogeneous components and whose morphisms are admissible grade transitions; adjoint pairs encoding typed round trips; and information-geometric interpretations in terms of KL gain, mirror descent with Bregman divergences, and Fisher natural gradients. Methodologically, we specify a utility--aware routing mechanism and objective that remain fully end-to-end differentiable. Analytic case studies and lightweight sanity checks illustrate selective morphic activation on hybrid symbolic-linguistic tasks. The framework unifies symbolic computation, geometry, and self--supervised learning within the \emph{graded transformer} formalism \cite{sh-89,sh-95}, while subsuming prior external-tool paradigms (e.g., Toolformer \cite{toolformer2023}) as a special case via functorial internalization.

📄 PDF Abstract BibTeX arXiv:2511.17840

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Structural Preservation and the Logical Expressiveness of Graph Neural Networks

2026-06-16 · Przemysław Andrzej Wałęga, Bernardo Cuenca Grau arxiv

Bridges between graph neural networks (GNNs) and logical formalisms have been established by fixing architectural choices, such as the types of aggregation, combination, and activation functions. These choices define res…

Topology-Informed Graph Transformer

2024-02-03 · Yun Young Choi, Sun Woo Park, Minho Lee, Youngho Woo

Transformers have revolutionized performance in Natural Language Processing and Vision, paving the way for their integration with Graph Neural Networks (GNNs). One key challenge in enhancing graph transformers is strengt…

Graph ClassificationGraph RegressionInductive BiasNode Classification

Pulling back symmetric Riemannian geometry for data analysis

2024-03-11 · Willem Diepeveen

Data sets tend to live in low-dimensional non-linear subspaces. Ideal data analysis tools for such data sets should therefore account for such non-linear geometry. The symmetric Riemannian geometry setting can be suitabl…

Internalizing LLM Reasoning via Discovery and Replay of Latent Actions

2026-02-04 · Zhenning Shi, Yijia Zhu, Junhan Shi, Xun Zhang 외 arxiv

The internalization of chain-of-thought processes into hidden states has emerged as a highly efficient paradigm for scaling test-time compute. However, existing activation steering methods rely on static control vectors …

Folding and Unfolding on Metagraphs

2020-12-03 · Ben Goertzel

Typed metagraphs are defined as hypergraphs with types assigned to hyperedges and their targets, and the potential to have targets of hyperedges connect to whole links as well as targets. Directed typed metagraphs (DTMGs…