paper-with-me

홈 › Papers

Data Fusion with Distributional Equivalence Test-then-pool

2026-03-12 · Linying Yang, Xing Liu, Robin J. Evans arxiv

Randomized controlled trials (RCTs) are the gold standard for causal inference, yet practical constraints often limit the size of the concurrent control arm. Borrowing control data from previous trials offers a potential efficiency gain, but naive borrowing can induce bias when historical and current populations differ. Existing test-then-pool (TTP) procedures address this concern by testing for equality of control outcomes between historical and concurrent trials before borrowing; however, standard implementations may suffer from reduced power or inadequate control of the Type-I error rate. We develop a new TTP framework that fuses control arms while rigorously controlling the Type-I error rate of the final treatment effect test. Our method employs kernel two-sample testing via maximum mean discrepancy (MMD) to capture distributional differences, and equivalence testing to avoid introducing uncontrolled bias, providing a more flexible and informative criterion for pooling. To ensure valid inference, we introduce partial bootstrap and partial permutation procedures for approximating null distributions in the presence of heterogeneous controls. We further establish the overall validity and consistency. We provide empirical studies demonstrating that the proposed approach achieves higher power than standard TTP methods while maintaining nominal error control, highlighting its value as a principled tool for leveraging historical controls in modern clinical trials.

📄 PDF Abstract BibTeX arXiv:2603.11867

Code (0)

등록된 구현이 없습니다.

Tasks

Two-sample testingCausal Inference

Similar Papers 제목 키워드 기반

Kernel Tests of Equivalence

2026-03-11 · Xing Liu, Axel Gandy arxiv

We propose novel kernel-based tests for assessing the equivalence between distributions. Traditional goodness-of-fit testing is inappropriate for concluding the absence of distributional differences, because failure to r…

Practical validation of synthetic pre-crash scenarios

2026-05-06 · Jian Wu, Ulrich Sander, Carol Flannagan, Jonas Bärgman arxiv

The representativeness of synthetic pre-crash scenarios is crucial for assessing the safety impact of Driving Automation Systems through virtual simulations. However, a gap remains in the robust evaluation of synthetic p…

Fuzzy Speculative Decoding for a Tunable Accuracy-Runtime Tradeoff

2025-02-28 · Maximilian Holsman, Yukun Huang, Bhuwan Dhingra

Speculative Decoding (SD) enforces strict distributional equivalence to the target model, limiting potential speed ups as distributions of near-equivalence achieve comparable outcomes in many cases. Furthermore, enforcin…

Reference-Specific Unlearning Metrics Can Hide the Truth: A Reality Check

2025-10-14 · Sungjun Cho, Dasol Hwang, Frederic Sala, Sangheum Hwang 외 arxiv

Current unlearning metrics for generative models evaluate success based on reference responses or classifier outputs rather than assessing the core objective: whether the unlearned model behaves indistinguishably from a …

Distributional Model Equivalence for Risk-Sensitive Reinforcement Learning

2023-07-04 · NeurIPS 2023 11 · Tyler Kastner, Murat A. Erdogdu, Amir-Massoud Farahmand

We consider the problem of learning models for risk-sensitive reinforcement learning. We theoretically demonstrate that proper value equivalence, a method of learning models which can be used to plan optimally in the ris…

Distributional Reinforcement Learningmodelreinforcement-learningReinforcement Learning