paper-with-me

홈 › Papers

Generalized Similarity U: A Non-parametric Test of Association Based on Similarity

2018-01-04 · Changshuai Wei, Qing Lu

Second generation sequencing technologies are being increasingly used for genetic association studies, where the main research interest is to identify sets of genetic variants that contribute to various phenotype. The phenotype can be univariate disease status, multivariate responses and even high-dimensional outcomes. Considering the genotype and phenotype as two complex objects, this also poses a general statistical problem of testing association between complex objects. We here proposed a similarity-based test, generalized similarity U (GSU), that can test the association between complex objects. We first studied the theoretical properties of the test in a general setting and then focused on the application of the test to sequencing association studies. Based on theoretical analysis, we proposed to use Laplacian kernel based similarity for GSU to boost power and enhance robustness. Through simulation, we found that GSU did have advantages over existing methods in terms of power and robustness. We further performed a whole genome sequencing (WGS) scan for Alzherimer Disease Neuroimaging Initiative (ADNI) data, identifying three genes, APOE, APOC1 and TOMM40, associated with imaging phenotype. We developed a C++ package for analysis of whole genome sequencing data using GSU. The source codes can be downloaded at https://github.com/changshuaiwei/gsu.

📄 PDF Abstract BibTeX arXiv:1801.01220

Code (1)

changshuaiwei/gsu 공식 구현

Similar Papers 제목 키워드 기반

Prior-Constrained Association Learning for Fine-Grained Generalized Category Discovery

2025-02-13 · Menglin Wang, Zhun Zhong, Xiaojin Gong

This paper addresses generalized category discovery (GCD), the task of clustering unlabeled data from potentially known or unknown categories with the help of labeled instances from each known category. Compared to tradi…

ClusteringRepresentation Learning

A Parametric Similarity Method: Comparative Experiments based on Semantically Annotated Large Datasets

2023-02-08 · Antonio De Nicola, Anna Formica, Michele Missikoff, Elaheh Pourabbas 외

We present the parametric method SemSimp aimed at measuring semantic similarity of digital resources. SemSimp is based on the notion of information content, and it leverages a reference ontology and taxonomic reasoning, …

Semantic SimilaritySemantic Textual Similarity

Fast Non-Parametric Tests of Relative Dependency and Similarity

2016-11-17 · Wacha Bounliphone, Eugene Belilovsky, Arthur Tenenhaus, Ioannis Antonoglou 외

We introduce two novel non-parametric statistical hypothesis tests. The first test, called the relative test of dependency, enables us to determine whether one source variable is significantly more dependent on a first t…

Data-Driven Nonparametric Existence and Association Problems

2017-11-22

We investigate two closely related nonparametric hypothesis testing problems. In the first problem (i.e., the existence problem), we test whether a testing data stream is generated by one of a set of composite distributi…

A Generalized Genetic Random Field Method for the Genetic Association Analysis of Sequencing Data

2025-08-18 · Ming Li, Zihuai He, Min Zhang, Xiaowei Zhan 외 arxiv

With the advance of high-throughput sequencing technologies, it has become feasible to investigate the influence of the entire spectrum of sequencing variations on complex human diseases. Although association studies uti…