paper-with-me

홈 › Papers

Panini: Continual Learning in Token Space via Structured Memory

2026-02-16 · Shreyas Rajesh, Pavan Holur, Mehmet Yigit Turali, Chenda Duan, Vwani Roychowdhury arxiv

Language models are increasingly used to reason over content they were not trained on, such as new documents, evolving knowledge, and user-specific data. A common approach is retrieval-augmented generation (RAG), which stores verbatim documents externally (as chunks) and retrieves only a relevant subset at inference time for an LLM to reason over. However, this results in inefficient usage of test-time compute (LLM repeatedly reasons over the same documents); moreover, chunk retrieval can inject irrelevant context that increases unsupported generation. We propose a human-like non-parametric continual learning framework, where the base model remains fixed, and learning occurs by integrating each new experience into an external semantic memory state that accumulates and consolidates itself continually. We present Panini, which realizes this by representing documents as Generative Semantic Workspaces (GSW) -- an entity- and event-aware network of question-answer (QA) pairs, sufficient for an LLM to reconstruct the experienced situations and mine latent knowledge via reasoning-grounded inference chains on the network. Given a query, Panini only traverses the continually-updated GSW (not the verbatim documents or chunks), and retrieves the most likely inference chains. Across six QA benchmarks, Panini achieves the highest average performance, 5%-7% higher than other competitive baselines, while using 2-30x fewer answer-context tokens, supports fully open-source pipelines, and reduces unsupported answers on curated unanswerable queries. The results show that efficient and accurate structuring of experiences at write time -- as achieved by the GSW framework -- yields both efficiency and reliability gains at read time. Code is available at https://github.com/roychowdhuryresearch/gsw-memory.

📄 PDF Abstract BibTeX arXiv:2602.15156

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

PaniniQA: Enhancing Patient Education Through Interactive Question Answering

2023-08-07 · Pengshan Cai, Zonghai Yao, Fei Liu, Dakuo Wang 외

Patient portal allows discharged patients to access their personalized discharge instructions in electronic health records (EHRs). However, many patients have difficulty understanding or memorizing their discharge instru…

Question Answering

Panini-Net: GAN Prior Based Degradation-Aware Feature Interpolation for Face Restoration

2022-03-16 · Yinhuai Wang, Yujie Hu, Jian Zhang

Emerging high-quality face restoration (FR) methods often utilize pre-trained GAN models (\textit{i.e.}, StyleGAN2) as GAN Prior. However, these methods usually struggle to balance realness and fidelity when facing vario…

Representation LearningSuper-Resolution

On Panini and the Generative Capacity of Contextualized Replacement Systems

2012-12-01 · COLING 2012 12 · Gerald Penn, Paul Kiparsky

HyperTokens: Controlling Token Dynamics for Continual Video-Language Understanding

2026-03-02 · Toan Nguyen, Yang Liu, Celso De Melo, Flora D. Salim arxiv

Continual VideoQA with multimodal LLMs is hindered by interference between tasks and the prohibitive cost of storing task-specific prompts. We introduce HyperTokens, a transformer-based token generator that produces fine…

Continual Fine-Tuning of Large Language Models via Program Memory

2026-05-13 · Hung Le, Svetha Venkatesh arxiv

Parameter-Efficient Fine-Tuning (PEFT), particularly Low-Rank Adaptation (LoRA), has become a standard approach for adapting Large Language Models (LLMs) under limited compute. However, in continual settings where models…

parameter-efficient fine-tuning