paper-with-me

홈 › Papers

IntLIM: Integration using Linear Models of metabolomics and gene expression data

2018-02-28

Integration of transcriptomic and metabolomic data improves functional interpretation of disease-related metabolomic phenotypes, and facilitates discovery of putative metabolite biomarkers and gene targets. For this reason, these data are increasingly collected in large cohorts, driving a need for the development of novel methods for their integration. Of note, clinical/translational studies typically provide snapshot gene and metabolite profiles and, oftentimes, most metabolites are not identified. Thus, in these types of studies, pathway/network approaches that take into account the complexity of gene-metabolite relationships may neither be applicable nor readily uncover novel relationships. With this in mind, we propose a simple linear modeling approach to capture phenotype-specific gene-metabolite associations, with the assumption that co-regulation patterns reflect functionally related genes and metabolites. The proposed linear model, metabolite ~ gene + phenotype + gene:phenotype, specifically evaluates whether gene-metabolite relationships differ by phenotype, by testing whether the relationship in one phenotype is significantly different from the relationship in another phenotype (via an interaction gene:phenotype p-value). Interaction p-values for all possible gene-metabolite pairs are computed and significant pairs are clustered by the directionality of associations. We implemented our approach as an R package, IntLIM, which includes a user-friendly Shiny app. We applied IntLIM to two published datasets, collected in NCI-60 cell lines and in human breast tumor and non-tumor tissue. We demonstrate that IntLIM captures relevant tumor-specific gene-metabolite associations involved in cancer-related pathways. and also uncover novel relationships that could be tested experimentally. The IntLIM R package is publicly available in GitHub (https://github.com/mathelab/IntLIM).

📄 PDF Abstract BibTeX arXiv:1802.10588

Code (1)

mathelab/IntLIM 공식 구현

Similar Papers 제목 키워드 기반

Multi-View Variational Autoencoder for Missing Value Imputation in Untargeted Metabolomics

2023-10-12 · Chen Zhao, Kuan-Jui Su, Chong Wu, Xuewei Cao 외

Background: Missing data is a common challenge in mass spectrometry-based metabolomics, which can lead to biased and incomplete analyses. The integration of whole-genome sequencing (WGS) data with metabolomics data has e…

Data IntegrationImputationMissing Values

Dynamical models for metabolomics data integration

2021-05-21 · Polina Lakrisenko, Daniel Weindl

As metabolomics datasets are becoming larger and more complex, there is an increasing need for model-based data integration and analysis to optimally leverage these data. Dynamical models of metabolism allow for the inte…

Data Integration

Scalable Randomized Kernel Methods for Multiview Data Integration and Prediction

2023-04-10 · Sandra E. Safo, Han Lu

We develop scalable randomized kernel methods for jointly associating data from multiple sources and simultaneously predicting an outcome or classifying a unit into one of two or more classes. The proposed methods model …

Data Integration

MetaboLLM: a metabolomics-specialized large language model for biochemical knowledge integration and predictive metabolite graph construction

2026-08-06 · Dohyun Ku, Min Gu Kwak, Francisco J. Pasquel, Jing Li arxiv

Metabolomics knowledge is distributed across heterogeneous resources and remains difficult to translate into predictive representations. We developed MetaboLLM, a metabolomics-specialized large language model adapted thr…

Continual Pretraining

MetaBench: A Multi-task Benchmark for Assessing LLMs in Metabolomics

2025-10-16 · Yuxing Lu, Xukai Zhao, J. Ben Tamo, Micky C. Nnamdi 외 arxiv

Large Language Models (LLMs) have demonstrated remarkable capabilities on general text; however, their proficiency in specialized scientific domains that require deep, interconnected knowledge remains largely uncharacter…

Text Generation