paper-with-me

Papers

On the Correlation between Random Variables and their Principal Components

2023-10-09 · Zenon Gniazdowski

The article attempts to find an algebraic formula describing the correlation coefficients between random variables and the principal components representing them. As a result of the analysis, starting from selected statistics relating to individual random variables, the equivalents of these statistics relating to a set of random variables were presented in the language of linear algebra, using the concepts of vector and matrix. This made it possible, in subsequent steps, to derive the expected formula. The formula found is identical to the formula used in Factor Analysis to calculate factor loadings. The discussion showed that it is possible to apply this formula to optimize the number of principal components in Principal Component Analysis, as well as to optimize the number of factors in Factor Analysis.

📄 PDF Abstract BibTeX arXiv:2310.06139

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On the clustering of correlated random variables

2019-09-07 · Zenon Gniazdowski, Dawid Kaliszewski

In this work, the possibility of clustering correlated random variables was examined, both because of their mutual similarity and because of their similarity to the principal components. The k-means algorithm and spectra…

ClusteringDiversity

Optimal whitening and decorrelation

2015-12-02 · Agnan Kessy, Alex Lewin, Korbinian Strimmer

Whitening, or sphering, is a common preprocessing step in statistical analysis to transform random variables to orthogonality. However, due to rotational freedom there are infinitely many possible whitening procedures. C…

An Information-theoretic Approach to Unsupervised Feature Selection for High-Dimensional Data

2019-10-08 · Shao-Lun Huang, Xiangxiang Xu, Lizhong Zheng

In this paper, we propose an information-theoretic approach to design the functional representations to extract the hidden common structure shared by a set of random variables. The main idea is to measure the common info…

feature selection

Agglomerative Info-Clustering

2017-01-18 · Chung Chan, Ali Al-Bashabsheh, Qiaoqiao Zhou

An agglomerative clustering of random variables is proposed, where clusters of random variables sharing the maximum amount of multivariate mutual information are merged successively to form larger clusters. Compared to t…

Clustering

Conditional canonical correlation estimation based on covariates with random forests

2020-11-23 · Cansu Alakus, Denis Larocque, Sebastien Jacquemont, Fanny Barlaam 외

Investigating the relationships between two sets of variables helps to understand their interactions and can be done with canonical correlation analysis (CCA). However, the correlation between the two sets can sometimes …

EEGElectroencephalogram (EEG)