paper-with-me

홈 › Papers

A Bayesian Approach to Low-Discrepancy Subset Selection

2026-02-16 · Nathan Kirk arxiv

Low-discrepancy designs play a central role in quasi-Monte Carlo methods and are increasingly influential in other domains such as machine learning, robotics and computer graphics, to name a few. In recent years, one such low-discrepancy construction method called subset selection has received a lot of attention. Given a large population, one optimally selects a small low-discrepancy subset with respect to a discrepancy-based objective. Versions of this problem are known to be NP-hard. In this text, we establish, for the first time, that the subset selection problem with respect to kernel discrepancies is also NP-hard. Motivated by this intractability, we propose a Bayesian Optimization procedure for the subset selection problem utilizing the recent notion of deep embedding kernels. We demonstrate the performance of the BO algorithm to minimize discrepancy measures and note that the framework is broadly applicable any design criteria.

📄 PDF Abstract BibTeX arXiv:2602.14607

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Optimizing Kernel Discrepancies via Subset Selection

2025-11-04 · Deyao Chen, François Clément, Carola Doerr, Nathan Kirk arxiv

Kernel discrepancies are a powerful tool for analyzing worst-case errors in quasi-Monte Carlo (QMC) methods. Building on recent advances in optimizing such discrepancy measures, we extend the subset selection problem to …

Bayesian subset selection and variable importance for interpretable prediction and classification

2021-04-20 · Daniel R. Kowal

Subset selection is a valuable tool for interpretable learning, scientific discovery, and data compression. However, classical subset selection is often avoided due to selection instability, lack of regularization, and d…

Data CompressionGeneral Classificationscientific discoveryUncertainty Quantification+1

Fair Bayesian Data Selection via Generalized Discrepancy Measures

2025-11-10 · Yixuan Zhang, Jiabin Luo, Zhenggang Wang, Feng Zhou 외 arxiv

Fairness concerns are increasingly critical as machine learning models are deployed in high-stakes applications. While existing fairness-aware methods typically intervene at the model level, they often suffer from high c…

Subset selection for linear mixed models

2021-07-27 · Daniel R. Kowal

Linear mixed models (LMMs) are instrumental for regression analysis with structured dependence, such as grouped, clustered, or multilevel data. However, selection among the covariates--while accounting for this structure…

Uncertainty Quantification

Bayesian information theoretic model-averaging stochastic item selection for computer adaptive testing: compromise-free item exposure

2025-04-22 · Joshua C. Chang, Edison Choe

The goal of Computer Adaptive Testing (CAT) is to reliably estimate an individual's ability as modeled by an item response theory (IRT) instrument using only a subset of the instrument's items. A secondary goal is to var…