paper-with-me

Papers

ProxyGuard: Direct Reliability Inference for Randomized Data Release Mechanisms with Shared Targets

2026-08-19 · Dipesh Tharu Mahato, Pramod Dhungana arxiv

Researchers often choose a proxy dataset from many releases, transformations, or seeds. Search can make an invalid release appear adequate, while one adequate release does not establish that its generator is reliable. ProxyGuard controls both errors using prespecified bounded risks and a sealed target set. Named-release mode corrects for multiplicity and certifies specific releases. Direct shared-target mode evaluates independent mechanism draws on a common target, lower-bounds their favorable-score rate, and subtracts a bound on favorable scores contributed by invalid releases. Conditional on the target, release scores are independent, yielding a finite-sample mechanism-reliability guarantee without independent target batches or assumptions on release-level $p$-value dependence. We show that the mean-only penalty is sharp and derive a smooth-score certificate with additive target concentration. In a registered three-requirement study, direct mode raises power from 5.6\% to 64.2\% at reliability 0.95, while named mode remains stronger under high-signal evidence. Prospective audits span full-pipeline Rice--TVAE, which retrains on every draw, and a non-tabular text mechanism.

📄 PDF Abstract BibTeX arXiv:2608.18643

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Directional Embedding Smoothing for Robust Vision Language Models

2026-03-16 · Ye Wang, Jing Liu, Toshiaki Koike-Akino arxiv

The safety and reliability of vision-language models (VLMs) are a crucial part of deploying trustworthy agentic AI systems. However, VLMs remain vulnerable to jailbreaking attacks that undermine their safety alignment to…

Private and Reliable Neural Network Inference

2022-10-27 · Nikola Jovanović, Marc Fischer, Samuel Steffen, Martin Vechev

Reliable neural networks (NNs) provide important inference-time reliability guarantees such as fairness and robustness. Complementarily, privacy-preserving NN inference protects the privacy of client data. So far these t…

FairnessPrivacy Preserving

Sample size planning for conditional counterfactual mean estimation with a K-armed randomized experiment

2024-03-06 · Gabriel Ruiz

We cover how to determine a sufficiently large sample size for a $K$-armed randomized experiment in order to estimate conditional counterfactual expectations in data-driven subgroups. The sub-groups can be output by any …

counterfactual

Estimating the Error of Randomized Newton Methods: A Bootstrap Approach

2020-01-01 · ICML 2020 1 · Miles Lopes, Jessie X.T. Chen

Randomized Newton methods have recently become the focus of intense research activity in large-scale and distributed optimization. Generally, these methods are based on a "computation-accuracy trade-off", which allows th…

Distributed Optimizationvalid

Causal inference with Machine Learning-Based Covariate Representation

2023-11-03 · Yuhang Wu, Jinghai He, Zeyu Zheng

Utilizing covariate information has been a powerful approach to improve the efficiency and accuracy for causal inference, which support massive amount of randomized experiments run on data-driven enterprises. However, st…

Causal Inference