paper-with-me

Papers

Meta-Analysis with Untrusted Data

2024-07-12 · Shiva Kaul, Geoffrey J. Gordon

[See paper for full abstract] Meta-analysis is a crucial tool for answering scientific questions. It is usually conducted on a relatively small amount of `trusted'' data -- ideally from randomized, controlled trials -- which allow causal effects to be reliably estimated with minimal assumptions. We show how to answer causal questions much more precisely by making two changes. First, we incorporate untrusted data drawn from large observational databases, related scientific literature and practical experience -- without sacrificing rigor or introducing strong assumptions. Second, we train richer models capable of handling heterogeneous trials, addressing a long-standing challenge in meta-analysis. Our approach is based on conformal prediction, which fundamentally produces rigorous prediction intervals, but doesn't handle indirect observations: in meta-analysis, we observe only noisy effects due to the limited number of participants in each trial. To handle noise, we develop a simple, efficient version of fully-conformal kernel ridge regression, based on a novel condition called idiocentricity. We introduce noise-correcting terms in the residuals and analyze their interaction with a `variance shaving'' technique. In multiple experiments on healthcare datasets, our algorithms deliver tighter, sounder intervals than traditional ones. This paper charts a new course for meta-analysis and evidence-based medicine, where heterogeneity and untrusted data are embraced for more nuanced and precise predictions.

📄 PDF Abstract BibTeX arXiv:2407.09387

Code (0)

등록된 구현이 없습니다.

Tasks

Conformal PredictionPrediction Intervals

Similar Papers 제목 키워드 기반

Document-Authored Control-Signal Impersonation: A Low-Cost Indirect Prompt Attack on RAG Safety Boundaries

2026-06-08 · Jianguo Zhu arxiv

Retrieval-augmented generation (RAG) systems often serialize user queries, retrieved documents, metadata, system labels, and task instructions into one natural-language prompt. We study a source-authority boundary failur…

Can a Multi-Hop Link Relying on Untrusted Amplify-and-Forward Relays Render Security?

2019-05-17 · Milad Tatar Mamaghani, Ali Kuhestani, Hamid Behroozi

Cooperative relaying is utilized as an efficient method for data communication in wireless sensor networks and the Internet of Things (IoT). However, sometimes due to the necessity of multi-hop relaying in such communica…

Decentralized Matrix Factorization with Heterogeneous Differential Privacy

2022-12-01 · Wentao Hu, Hui Fang

Conventional matrix factorization relies on centralized collection of users' data for recommendation, which might introduce an increased risk of privacy leakage especially when the recommender is untrusted. Existing diff…

Untrusted Content Masking for Web Agents with Security Guarantees

2026-07-06 · Kristina Nikolić, Egor Zverev, Javier Rando, Matthew Jagielski 외 arxiv

Defenses that provide security guarantees against prompt injection attacks rely on strict isolation between trusted instructions and untrusted data. In text-based environments such as tool-use APIs, this separation arise…

When can we trust untrusted monitoring? A safety case sketch across collusion strategies

2026-02-24 · Nelson Gardner-Challis, Jonathan Bostock, Georgiy Kozhevnikov, Morgan Sinclaire 외 arxiv

AIs are increasingly being deployed with greater autonomy and capabilities, which increases the risk that a misaligned AI may be able to cause catastrophic harm. Untrusted monitoring -- using one untrusted model to overs…