paper-with-me

홈 › Papers

WildCat: Near-Linear Attention in Theory and Practice

2026-02-10 · Tobias Schröder, Lester Mackey arxiv

We introduce WildCat, a high-accuracy, low-cost approach to compressing the attention mechanism in neural networks. While attention is a staple of modern network architectures, it is also notoriously expensive to deploy due to resource requirements that scale quadratically with the input sequence length $n$. WildCat avoids these quadratic costs by only attending over a small weighted coreset. Crucially, we select the coreset using a fast but spectrally-accurate subsampling algorithm -- randomly pivoted Cholesky -- and weight the elements optimally to minimise reconstruction error. Remarkably, given bounded inputs, WildCat approximates exact attention with super-polynomial $O(n^{-\sqrt{\log(\log(n))}})$ error decay while running in near-linear $O(n^{1+o(1)})$ time. In contrast, prior practical approximations either lack error guarantees or require quadratic runtime to guarantee such high fidelity. We couple this advance with a GPU-optimized PyTorch implementation and a suite of benchmark experiments demonstrating the benefits of WildCat for image generation, image classification, and language model KV cache compression.

📄 PDF Abstract BibTeX arXiv:2602.10056

Code (0)

등록된 구현이 없습니다.

Tasks

Image ClassificationImage Generation

Similar Papers 제목 키워드 기반

Application of deep learning to camera trap data for ecologists in planning / engineering -- Can captivity imagery train a model which generalises to the wild?

2021-11-24 · Ryan Curry, Cameron Trotter, Andrew Stephen McGough

Understanding the abundance of a species is the first step towards understanding both its long-term sustainability and the impact that we may be having upon it. Ecologists use camera traps to remotely survey for the pres…

image-classificationImage ClassificationImage ManipulationImage Segmentation+3

WILDCAT: Weakly Supervised Learning of Deep ConvNets for Image Classification, Pointwise Localization and Segmentation

2017-07-01 · CVPR 2017 7 · Thibaut Durand, Taylor Mordan, Nicolas Thome, Matthieu Cord

This paper introduces WILDCAT, a deep learning method which jointly aims at aligning image regions for gaining spatial invariance and learning strongly localized features. Our model is trained using only global image lab…

General Classificationimage-classificationImage ClassificationObject Localization+4

WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild

2025-06-16 · Morris Alper, David Novotny, Filippos Kokkinos, Hadar Averbuch-Elor 외

Despite recent advances in sparse novel view synthesis (NVS) applied to object-centric scenes, scene-level NVS remains a challenge. A central issue is the lack of available clean multi-view training data, beyond manually…

Novel View Synthesis

ENA: Efficient N-dimensional Attention

2025-08-16 · Yibo Zhong arxiv

Efficient modeling of long sequences of high-order data requires a more efficient architecture than Transformer. In this paper, we investigate two key aspects of extending linear recurrent models, especially those origin…

Preconditioned DeltaNet: Curvature-aware Sequence Modeling for Linear Recurrences

2026-04-22 · Neehal Tumma, Noel Loo, Daniela Rus arxiv

To address the increasing long-context compute limitations of softmax attention, several subquadratic recurrent operators have been developed. This work includes models such as Mamba-2, DeltaNet, Gated DeltaNet (GDN), an…