paper-with-me

Papers

Model-based clustering for identifying disease-associated SNPs in case-control genome-wide association studies

2018-06-21 · Yan Xu, Li Xing, Jessica Su, Xuekui Zhang, Weiliang Qiu

Genome-wide association studies (GWASs) aim to detect genetic risk factors for complex human diseases by identifying disease-associated single-nucleotide polymorphisms (SNPs). The traditional SNP-wise approach along with multiple testing adjustment is over-conservative and lack of power in many GWASs. In this article, we proposed a model-based clustering method that transforms the challenging high-dimension-small-sample-size problem to low-dimension-large-sample-size problem and borrows information across SNPs by grouping SNPs into three clusters. We pre-specify the patterns of clusters by minor allele frequencies of SNPs between cases and controls, and enforce the patterns with prior distributions. In the simulation studies our proposed novel model outperform traditional SNP-wise approach by showing better controls of false discovery rate (FDR) and higher sensitivity. We re-analyzed two real studies to identifying SNPs associated with severe bortezomib-induced peripheral neuropathy (BiPN) in patients with multiple myeloma (MM). The original analysis in the literature failed to identify SNPs after FDR adjustment. Our proposed method not only detected the reported SNPs after FDR adjustment but also discovered a novel BiPN-associated SNP rs4351714 that has been reported to be related to MM in another study.

📄 PDF Abstract BibTeX arXiv:1806.08456

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Similar Papers 제목 키워드 기반

Identifying Genetic Risk Factors via Sparse Group Lasso with Group Graph Structure

2017-09-12 · Tao Yang, Paul Thompson, Sihai Zhao, Jieping Ye

Genome-wide association studies (GWA studies or GWAS) investigate the relationships between genetic variants such as single-nucleotide polymorphisms (SNPs) and individual traits. Recently, incorporating biological priors…

Variable Selection

Assessing the Reproducibility of Machine-learning-based Biomarker Discovery in Parkinson's Disease

2023-04-06 · Ali Amelia, Lourdes Pena-Castillo, Hamid Usefi

Genome-Wide Association Studies (GWAS) help identify genetic variations in people with diseases such as Parkinson's disease (PD), which are less common in those without the disease. Thus, GWAS data can be used to identif…

Data Integrationfeature selection

Predicting Pathogenicity Of nsSNPs Associated With Rb1 -- An In Silico Approach

2023-07-16 · Anum Munir

Single nucleotide polymorphisms (SNPs) are variations at specific locations in DNA. Sequence responsible for marking genes associated with diseases or tracking inherited diseases within The family. These variations in th…

FANCA: In-Silico deleterious mutation analysis for early prediction of leukemia

2021-07-19 · Madiha Hameed, Abdul Majiid, Asifullah Khan

As a novel biomarker from the Fanconi anemia complementation group (FANC) family, FANCA is antigens to Leukemia cancer. The overexpression of FANCA has predicted the second most common cancer in the world that is respons…

Drug Discovery

DuAL-Net: A Hybrid Framework for Alzheimer's Disease Prediction from Whole-Genome Sequencing via Local SNP Windows and Global Annotations

2025-05-31 · Eun Hye Lee, Taeho Jo

Alzheimer's disease (AD) dementia is the most common form of dementia. With the emergence of disease-modifying therapies, predicting disease risk before symptom onset has become critical. We introduce DuAL-Net, a hybrid …

Disease Prediction