paper-with-me

홈 › Papers

A Unified Probabilistic Framework for Dictionary Learning with Parsimonious Activation

2025-09-30 · Zihui Zhao, Yuanbo Tang, Jieyu Ren, Xiaoping Zhang, Yang Li arxiv

Dictionary learning is traditionally formulated as an $L_1$-regularized signal reconstruction problem. While recent developments have incorporated discriminative, hierarchical, or generative structures, most approaches rely on encouraging representation sparsity over individual samples that overlook how atoms are shared across samples, resulting in redundant and sub-optimal dictionaries. We introduce a parsimony promoting regularizer based on the row-wise $L_\infty$ norm of the coefficient matrix. This additional penalty encourages entire rows of the coefficient matrix to vanish, thereby reducing the number of dictionary atoms activated across the dataset. We derive the formulation from a probabilistic model with Beta-Bernoulli priors, which provides a Bayesian interpretation linking the regularization parameters to prior distributions. We further establish theoretical calculation for optimal hyperparameter selection and connect our formulation to both Minimum Description Length, Bayesian model selection and pathlet learning. Extensive experiments on benchmark datasets demonstrate that our method achieves substantially improved reconstruction quality (with a 20\% reduction in RMSE) and enhanced representation sparsity, utilizing fewer than one-tenth of the available dictionary atoms, while empirically validating our theoretical analysis.

📄 PDF Abstract BibTeX arXiv:2509.25690

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the Sparsity-Storage-Accuracy Tradeoff in Parsimoniously Activated Dictionary Learning

2026-06-21 · Zihui Zhao, Yuanbo Tang, Yang Li arxiv

Dictionary learning has long been studied from both optimization and probabilistic perspectives. While formulations with element-wise sparsity regularization (e.g., L1-based sparse coding) admit well-established probabil…

PaCE: Parsimonious Concept Engineering for Large Language Models

2024-06-06 · Jinqi Luo, Tianjiao Ding, Kwan Ho Ryan Chan, Darshan Thaker 외

Large Language Models (LLMs) are being used for a wide variety of tasks. While they are capable of generating human-like responses, they can also produce undesirable output including potentially harmful information, raci…

Prompt Engineering

Closed-form Marginal Likelihood in Gamma-Poisson Matrix Factorization

2018-01-05 · ICML 2018 7 · Louis Filstroff, Alberto Lumbreras, Cédric Févotte

We present novel understandings of the Gamma-Poisson (GaP) model, a probabilistic matrix factorization model for count data. We show that GaP can be rewritten free of the score/activation matrix. This gives us new insigh…

Form

Group Iterative Spectrum Thresholding for Super-Resolution Sparse Spectral Selection

2012-07-28 · Yiyuan She, Huanghuang Li, Jiangping Wang, Dapeng Wu

Recently, sparsity-based algorithms are proposed for super-resolution spectrum estimation. However, to achieve adequately high resolution in real-world signal analysis, the dictionary atoms have to be close to each other…

compressed sensingSuper-Resolution

Dictionary Learning Improves Patch-Free Circuit Discovery in Mechanistic Interpretability: A Case Study on Othello-GPT

2024-02-19 · Zhengfu He, Xuyang Ge, Qiong Tang, Tianxiang Sun 외

Sparse dictionary learning has been a rapidly growing technique in mechanistic interpretability to attack superposition and extract more human-understandable features from model activations. We ask a further question bas…

Dictionary Learning