paper-with-me

Papers

Group Representational Position Encoding

2025-12-08 · Yifan Zhang, Zixiang Chen, Yifeng Liu, Zhen Qin, Huizhuo Yuan, Kangping Xu, Yang Yuan, Quanquan Gu, Andrew Chi-Chih Yao arxiv

We present GRAPE (Group Representational Position Encoding), a unified framework for positional encoding based on group actions. GRAPE unifies two families of mechanisms: (i) multiplicative rotations (Multiplicative GRAPE) in $\operatorname{SO}(d)$ and (ii) additive logit biases (Additive GRAPE) arising from unipotent actions in the general linear group $\mathrm{GL}$. In Multiplicative GRAPE, a position $n \in \mathbb{Z}$ (or $t \in \mathbb{R}$) acts as $\mathbf{G}(n) = \exp(n \, ω\, \mathbf{L})$ with a rank-2 skew-symmetric generator $\mathbf{L} \in \mathbb{R}^{d \times d}$, yielding a relative, compositional, norm-preserving map with a closed-form matrix exponential. RoPE is recovered exactly when the $d/2$ planes correspond to canonical coordinate pairs with a log-uniform spectrum. Learned commuting subspaces and compact non-commuting mixtures strictly extend this geometry to capture cross-subspace feature coupling at $O(d)$ and $O(r d)$ cost per head, respectively. In Additive GRAPE, additive logits arise from rank-1 (or low-rank) unipotent actions, recovering ALiBi and the Forgetting Transformer (FoX) as exact special cases while preserving an exact relative law and streaming cacheability. Overall, GRAPE provides a principled design space for positional geometry in long-context models, subsuming RoPE and ALiBi as special cases. Project page: https://github.com/model-architectures/GRAPE.

📄 PDF Abstract BibTeX arXiv:2512.07805

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Group Equivariant Stand-Alone Self-Attention For Vision

2020-10-02 · ICLR 2021 1 · David W. Romero, Jean-Baptiste Cordonnier

We provide a general self-attention formulation to impose group equivariance to arbitrary symmetry groups. This is achieved by defining positional encodings that are invariant to the action of the group considered. Since…

Improving Equivariant Networks with Probabilistic Symmetry Breaking

2025-03-27 · Hannah Lawrence, Vasco Portilheiro, Yan Zhang, Sékou-Oumar Kaba

Equivariance encodes known symmetries into neural networks, often enhancing generalization. However, equivariant networks cannot break symmetries: the output of an equivariant network must, by definition, have at least t…

Generalization BoundsInductive Bias

Polynomial Implicit Neural Representations For Large Diverse Datasets

2023-03-20 · CVPR 2023 1 · Rajhans Singh, Ankita Shukla, Pavan Turaga

Implicit neural representations (INR) have gained significant popularity for signal and image representation for many end-tasks, such as superresolution, 3D modeling, and more. Most INR architectures rely on sinusoidal p…

Conditional Image GenerationImage Generation

PINs: Progressive Implicit Networks for Multi-Scale Neural Representations

2022-02-09 · Zoe Landgraf, Alexander Sorkine Hornung, Ricardo Silveira Cabral

Multi-layer perceptrons (MLP) have proven to be effective scene encoders when combined with higher-dimensional projections of the input, commonly referred to as \textit{positional encoding}. However, scenes with a wide f…

Distributed neural encoding of binding to thematic roles

2021-10-24 · Matthias Lalisse, Paul Smolensky

A framework and method are proposed for the study of constituent composition in fMRI. The method produces estimates of neural patterns encoding complex linguistic structures, under the assumption that the contributions o…

Sentence