paper-with-me

Papers

Multi-level hypothesis testing for populations of heterogeneous networks

2018-09-07 · Guilherme Gomes, Vinayak Rao, Jennifer Neville

In this work, we consider hypothesis testing and anomaly detection on datasets where each observation is a weighted network. Examples of such data include brain connectivity networks from fMRI flow data, or word co-occurrence counts for populations of individuals. Current approaches to hypothesis testing for weighted networks typically requires thresholding the edge-weights, to transform the data to binary networks. This results in a loss of information, and outcomes are sensitivity to choice of threshold levels. Our work avoids this, and we consider weighted-graph observations in two situations, 1) where each graph belongs to one of two populations, and 2) where entities belong to one of two populations, with each entity possessing multiple graphs (indexed e.g. by time). Specifically, we propose a hierarchical Bayesian hypothesis testing framework that models each population with a mixture of latent space models for weighted networks, and then tests populations of networks for differences in distribution over components. Our framework is capable of population-level, entity-specific, as well as edge-specific hypothesis testing. We apply it to synthetic data and three real-world datasets: two social media datasets involving word co-occurrences from discussions on Twitter of the political unrest in Brazil, and on Instagram concerning Attention Deficit Hyperactivity Disorder (ADHD) and depression drugs, and one medical dataset involving fMRI brain-scans of human subjects. The results show that our proposed method has lower Type I error and higher statistical power compared to alternatives that need to threshold the edge weights. Moreover, they show our proposed method is better suited to deal with highly heterogeneous datasets.

📄 PDF Abstract BibTeX arXiv:1809.02512

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly DetectionTwo-sample testing

Similar Papers 제목 키워드 기반

A model of multiple hypothesis testing

2021-04-27 · Davide Viviano, Kaspar Wuthrich, Paul Niehaus

Multiple hypothesis testing practices vary widely, without consensus on which are appropriate when. This paper provides an economic foundation for these practices designed to capture leading examples, such as regulatory …

model

A Sampling-based Framework for Hypothesis Testing on Large Attributed Graphs

2024-03-20 · Yun Wang, Chrysanthi Kosyfaki, Sihem Amer-Yahia, Reynold Cheng

Hypothesis testing is a statistical method used to draw conclusions about populations from sample data, typically represented in tables. With the prevalence of graph representations in real-life applications, hypothesis …

Graph Sampling

A Distance Correlation-based Kernel for Nonlinear Causal Clustering in Heterogeneous Populations

2021-05-21 · NeurIPS 2021 12 · Alex Markham, Moritz Grosse-Wentrup

We consider the problem of causal structure learning in the setting of heterogeneous populations, i.e., populations in which a single causal structure does not adequately represent all population members, as is common in…

Clustering

Hypothesis Testing for Quantifying LLM-Human Misalignment in Multiple Choice Settings

2025-06-17 · Harbin Hong, Sebastian Caldas, Liu Leqi

As Large Language Models (LLMs) increasingly appear in social science research (e.g., economics and marketing), it becomes crucial to assess how well these models replicate human behavior. In this work, using hypothesis …

Decision MakingLanguage ModelingLanguage ModellingMarketing+1

Heterogeneous Dense Subhypergraph Detection

2021-04-08 · Mingao Yuan, Zuofeng Shang

We study the problem of testing the existence of a heterogeneous dense subhypergraph. The null hypothesis corresponds to a heterogeneous Erd\"{o}s-R\'{e}nyi uniform random hypergraph and the alternative hypothesis corres…