paper-with-me

Papers

Distance Correlation Methods for Discovering Associations in Large Astrophysical Databases

2013-08-19 · Elizabeth Martinez-Gomez, Mercedes T. Richards, Donald St. P. Richards

High-dimensional, large-sample astrophysical databases of galaxy clusters, such as the Chandra Deep Field South COMBO-17 database, provide measurements on many variables for thousands of galaxies and a range of redshifts. Current understanding of galaxy formation and evolution rests sensitively on relationships between different astrophysical variables; hence an ability to detect and verify associations or correlations between variables is important in astrophysical research. In this paper, we apply a recently defined statistical measure called the distance correlation coefficient which can be used to identify new associations and correlations between astrophysical variables. The distance correlation coefficient applies to variables of any dimension; it can be used to determine smaller sets of variables that provide equivalent astrophysical information; it is zero only when variables are independent; and it is capable of detecting nonlinear associations that are undetectable by the classical Pearson correlation coefficient. Hence, the distance correlation coefficient provides more information than the Pearson coefficient. We analyze numerous pairs of variables in the COMBO-17 database with the distance correlation method and with the maximal information coefficient. We show that the Pearson coefficient can be estimated with higher accuracy from the corresponding distance correlation coefficient than from the maximal information coefficient. For given values of the Pearson coefficient, the distance correlation method has a greater ability than the maximal information coefficient to resolve astrophysical data into highly concentrated V-shapes, which enhances classification and pattern identification. These results are observed over a range of redshifts beyond the local universe and for galaxies from elliptical to spiral.

📄 PDF Abstract BibTeX arXiv:1308.3925

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Discovering Association with Copula Entropy

2019-07-29 · Jian Ma

Discovering associations is of central importance in scientific practices. Currently, most researches consider only linear association measured by correlation coefficient, which has its theoretical limitations. In this p…

Non-linear correlation analysis in financial markets using hierarchical clustering

2023-01-12 · J. E. Salgado-Hernández, Manan Vyas

Distance correlation coefficient (DCC) can be used to identify new associations and correlations between multiple variables. The distance correlation coefficient applies to variables of any dimension, can be used to dete…

Clustering

Sparse Canonical Correlation Analysis via Concave Minimization

2019-09-17 · Omid S. Solari, James B. Brown, Peter J. Bickel

A new approach to the sparse Canonical Correlation Analysis (sCCA)is proposed with the aim of discovering interpretable associations in very high-dimensional multi-view, i.e.observations of multiple sets of variables on …

Computational Efficiency

GloTSFormer: Global Video Text Spotting Transformer

2024-01-08 · Han Wang, Yanjie Wang, Yang Li, Can Huang

Video Text Spotting (VTS) is a fundamental visual task that aims to predict the trajectories and content of texts in a video. Previous works usually conduct local associations and apply IoU-based distance and complex pos…

Text Spotting

Comprehensive Metapath-based Heterogeneous Graph Transformer for Gene-Disease Association Prediction

2025-01-14 · Wentao Cui, Shoubo Li, Chen Fang, Qingqing Long 외

Discovering gene-disease associations is crucial for understanding disease mechanisms, yet identifying these associations remains challenging due to the time and cost of biological experiments. Computational methods are …