paper-with-me

Papers

LLM Sparsity Prior for Robust Feature Selection

2026-05-21 · Caleb Skinner, Yihan Guo, Meng Li arxiv

Large language models (LLMs) offer a scalable mechanism to elicit domain-informed prior information for high-dimensional variable selection. However, existing methods such as LLM-Lasso are sensitive to weight quality, with performance degrading substantially when LLM-generated weights are inaccurate. To address this challenge, we first introduce a framework for quantifying the quality of LLM-generated weights, enabling rigorous evaluation of LLM-informed methods across varying weight regimes. We then propose the LLM Sparsity Prior (LSP), which integrates LLM-generated weights into the prior inclusion probabilities of Spike-and-Slab and Spike-and-Slab Lasso models via two interpretable hyperparameters governing global sparsity and weight concentration. Hierarchical hyperpriors on these parameters allow the model to dynamically discount uninformative or misleading weights, improving robustness without sacrificing gains when weights are accurate. Finally, we develop principled prompt engineering strategies and validate the method on a private medical dataset studying Acute Kidney Injury. LSP improves prediction accuracy and identifies clinically relevant features missed by the baselines, with robustness to prompt variation and particular effectiveness in low-data regimes.

📄 PDF Abstract BibTeX arXiv:2605.23102

Code (0)

등록된 구현이 없습니다.

Tasks

Prompt Engineering

Similar Papers 제목 키워드 기반

The Illusion of the Illusion of Sparsity: An exercise in prior sensitivity

2020-09-29 · Bruno Fava, Hedibert F. Lopes

The emergence of Big Data raises the question of how to model economic relations when there is a large number of possible explanatory variables. We revisit the issue by comparing the possibility of using dense or sparse …

SensitivityVariable Selection

Selection Plateau and a Sparsity-Dependent Hierarchy of Pruning Features

2026-05-10 · Guangqi Li, Yongxin Li arxiv

We identify a Selection Plateau phenomenon in one-shot neural network pruning: all rank-monotone weight scorers converge to identical accuracy at fixed sparsity, independent of functional form. We propose the Sparsity-In…

Network Pruning

Causally-Guided Diffusion for Stable Feature Selection

2026-03-21 · Arun Vignesh Malarkkan, Xinyuan Wang, Kunpeng Liu, Denghui Zhang 외 arxiv

Feature selection is fundamental to robust data-centric AI, but most existing methods optimize predictive performance under a single data distribution. This often selects spurious features that fail under distribution sh…

Probabilistic Multi-Task Feature Selection

2010-12-01 · NeurIPS 2010 12 · Yu Zhang, Dit-yan Yeung, Qian Xu

Recently, some variants of the $l_1$ norm, particularly matrix norms such as the $l_{1,2}$ and $l_{1,\infty}$ norms, have been widely used in multi-task learning, compressed sensing and other related areas to enforce spa…

Cancer Classificationcompressed sensingfeature selectionMulti-Task Learning

Sparsity-based Feature Selection for Anomalous Subgroup Discovery

2022-01-06 · Girmaw Abebe Tadesse, William Ogallo, Catherine Wanjiru, Charles Wachira 외

Anomalous pattern detection aims to identify instances where deviation from normalcy is evident, and is widely applicable across domains. Multiple anomalous detection techniques have been proposed in the state of the art…

feature selectionSubgroup Discovery