paper-with-me

Papers

Restricted Block Permutation for Two-Sample Testing

2025-11-29 · Jungwoo Ho arxiv

We study a structured permutation scheme for two-sample testing that restricts permutations to single cross-swaps between block-selected representatives. Our analysis yields three main results. First, we provide an exact validity construction that applies to any fixed restricted permutation set. Second, for both the difference of sample means and the unbiased $\widehat{\mathrm{MMD}}^{2}$ estimator, we derive closed-form one-swap increment identities whose conditional variances scale as $O(h^{2})$, in contrast to the $Θ(h)$ increment variability under full relabeling. This increment-level variance contraction sharpens the Bernstein--Freedman variance proxy and leads to substantially smaller permutation critical values. Third, we obtain explicit, data-dependent expressions for the resulting critical values and statistical power. Together, these results show that block-restricted one-swap permutations can achieve strictly higher power than classical full permutation tests while maintaining exact finite-sample validity, without relying on pessimistic worst-case Lipschitz bounds.

📄 PDF Abstract BibTeX arXiv:2512.00668

Code (0)

등록된 구현이 없습니다.

Tasks

Two-sample testing

Similar Papers 제목 키워드 기반

r-local sensing: Improved algorithm and applications

2021-10-26 · Ahmed Ali Abbasi, Abiy Tasissa, Shuchin Aeron

The unlabeled sensing problem is to solve a noisy linear system of equations under unknown permutation of the measurements. We study a particular case of the problem where the permutations are restricted to be r-local, i…

Cheap Permutation Testing

2025-02-11 · Carles Domingo-Enrich, Raaz Dwivedi, Lester Mackey

Permutation tests are a popular choice for distinguishing distributions and testing independence, due to their exact, finite-sample control of false positives and their minimax optimality when paired with U-statistics. H…

Block-Value Symmetries in Probabilistic Graphical Models

2018-07-02 · Gagan Madan, Ankit Anand, Mausam, Parag Singla

One popular way for lifted inference in probabilistic graphical models is to first merge symmetric states into a single cluster (orbit) and then use these for downstream inference, via variations of orbital MCMC [Niepert…

The Chi-Square Test of Distance Correlation

2019-12-27 · Cencheng Shen, Sambit Panda, Joshua T. Vogelstein

Distance correlation has gained much recent attention in the data science community: the sample statistic is straightforward to compute and asymptotically equals zero if and only if independence, making it an ideal choic…

valid

Trustworthy Feature Importance Avoids Unrestricted Permutations

2026-04-13 · Emanuele Borgonovo, Francesco Cappelli, Xuefei Lu, Elmar Plischke 외 arxiv

Feature importance methods using unrestricted permutations are flawed due to extrapolation errors; such errors appear in all non-trivial variable importance approaches. We propose three new approaches: conditional model …

Feature Importance