paper-with-me

Papers

Network Cluster-Robust Inference

2021-03-02 · Michael P. Leung

Since network data commonly consists of observations from a single large network, researchers often partition the network into clusters in order to apply cluster-robust inference methods. Existing such methods require clusters to be asymptotically independent. Under mild conditions, we prove that, for this requirement to hold for network-dependent data, it is necessary and sufficient that clusters have low conductance, the ratio of edge boundary size to volume. This yields a simple measure of cluster quality. We find in simulations that when clusters have low conductance, cluster-robust methods control size better than HAC estimators. However, for important classes of networks lacking low-conductance clusters, the former can exhibit substantial size distortion. To determine the number of low-conductance clusters and construct them, we draw on results in spectral graph theory that connect conductance to the spectrum of the graph Laplacian. Based on these results, we propose to use the spectrum to determine the number of low-conductance clusters and spectral clustering to construct them.

📄 PDF Abstract BibTeX arXiv:2103.01470

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Methods 이 논문이 사용한 방법론

Spectral Clustering Spectral clustering has attracted increasing attention due to the promising ability in dealing with nonlinearly separable datasets [15], [16]. In spectral clustering, the…

Similar Papers 제목 키워드 기반

Improved Inference for CSDID Using the Cluster Jackknife

2026-02-12 · Sunny R. Karim, Morten Ørregaard Nielsen, James G. MacKinnon, Matthew D. Webb arxiv

Obtaining reliable inferences with traditional difference-in-differences (DiD) methods can be difficult. Problems can arise when both outcomes and errors are serially correlated, when there are few clusters or few treate…

Testing for the appropriate level of clustering in linear regression models

2023-01-11 · James G. MacKinnon, Morten Ørregaard Nielsen, Matthew D. Webb

The overwhelming majority of empirical research that uses cluster-robust inference assumes that the clustering structure is known, even though there are often several possible ways in which a dataset could be clustered. …

Clusteringregression

Evaluating the statistical significance of biclusters

2015-12-01 · NeurIPS 2015 12 · Jason D. Lee, Yuekai Sun, Jonathan E. Taylor

Biclustering (also known as submatrix localization) is a problem of high practical relevance in exploratory analysis of high-dimensional data. We develop a framework for performing statistical inference on biclusters fou…

Asymptotic Theory for Two-Way Clustering

2023-01-10 · Luther Yap

This paper proves a new central limit theorem for a sample that exhibits two-way dependence and heterogeneity across clusters. Statistical inference for situations with both two-way dependence and cluster heterogeneity h…

Clusteringregressionvalid

Inference with K-means

2024-10-04 · Alfred K. Adzika, Prudence Djagba

This thesis aims to invent new approaches for making inferences with the k-means algorithm. k-means is an iterative clustering algorithm that randomly assigns k centroids, then assigns data points to the nearest centroid…

ClusteringDensity Estimation