paper-with-me

홈 › Papers

SIMPLE: Statistical Inference on Membership Profiles in Large Networks

2019-10-03 · Jianqing Fan, Yingying Fan, Xiao Han, Jinchi Lv

Network data is prevalent in many contemporary big data applications in which a common interest is to unveil important latent links between different pairs of nodes. Yet a simple fundamental question of how to precisely quantify the statistical uncertainty associated with the identification of latent links still remains largely unexplored. In this paper, we propose the method of statistical inference on membership profiles in large networks (SIMPLE) in the setting of degree-corrected mixed membership model, where the null hypothesis assumes that the pair of nodes share the same profile of community memberships. In the simpler case of no degree heterogeneity, the model reduces to the mixed membership model for which an alternative more robust test is also proposed. Both tests are of the Hotelling-type statistics based on the rows of empirical eigenvectors or their ratios, whose asymptotic covariance matrices are very challenging to derive and estimate. Nevertheless, their analytical expressions are unveiled and the unknown covariance matrices are consistently estimated. Under some mild regularity conditions, we establish the exact limiting distributions of the two forms of SIMPLE test statistics under the null hypothesis and contiguous alternative hypothesis. They are the chi-square distributions and the noncentral chi-square distributions, respectively, with degrees of freedom depending on whether the degrees are corrected or not. We also address the important issue of estimating the unknown number of communities and establish the asymptotic properties of the associated test statistics. The advantages and practical utility of our new procedures in terms of both size and power are demonstrated through several simulation examples and real network applications.

📄 PDF Abstract BibTeX arXiv:1910.01734

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

SIMPLE-RC: Group Network Inference with Non-Sharp Nulls and Weak Signals

2022-10-31 · Jianqing Fan, Yingying Fan, Jinchi Lv, Fan Yang

Large-scale network inference with uncertainty quantification has important applications in natural, social, and medical sciences. The recent work of Fan, Fan, Han and Lv (2022) introduced a general framework of statisti…

Uncertainty Quantification

Fundamental Limits of Membership Inference Attacks on Machine Learning Models

2023-10-20 · Eric Aubinais, Elisabeth Gassiat, Pablo Piantanida

Membership inference attacks (MIA) can reveal whether a particular data point was part of the training dataset, potentially exposing sensitive information about individuals. This article provides theoretical guarantees b…

Diversity

Has My System Prompt Been Used? Large Language Model Prompt Membership Inference

2025-02-14 · Roman Levin, Valeriia Cherepanova, Abhimanyu Hans, Avi Schwarzschild 외

Prompt engineering has emerged as a powerful technique for optimizing large language models (LLMs) for specific applications, enabling faster prototyping and improved performance, and giving rise to the interest of the c…

Language ModelingLanguage ModellingLarge Language ModelPrompt Engineering

Canary in a Coalmine: Better Membership Inference with Ensembled Adversarial Queries

2022-10-19 · Yuxin Wen, Arpit Bansal, Hamid Kazemi, Eitan Borgnia 외

As industrial applications are increasingly automated by machine learning models, enforcing personal data ownership and intellectual property rights requires tracing training data back to their rightful owners. Membershi…

Context-Aware Membership Inference Attacks against Pre-trained Large Language Models

2024-09-11 · Hongyan Chang, Ali Shahin Shamsabadi, Kleomenis Katevas, Hamed Haddadi 외

Prior Membership Inference Attacks (MIAs) on pre-trained Large Language Models (LLMs), adapted from classification model attacks, fail due to ignoring the generative process of LLMs across token sequences. In this paper,…

Memorization