paper-with-me

홈 › Papers

Probabilistic Approach for Evaluating Metabolite Sample Integrity

2015-06-14

The success of metabolomics studies depends upon the "fitness" of each biological sample used for analysis: it is critical that metabolite levels reported for a biological sample represent an accurate snapshot of the studied organism's metabolite profile at time of sample collection. Numerous factors may compromise metabolite sample fitness, including chemical and biological factors which intervene during sample collection, handling, storage, and preparation for analysis. We propose a probabilistic model for the quantitative assessment of metabolite sample fitness. Collection and processing of nuclear magnetic resonance (NMR) and ultra-performance liquid chromatography (UPLC-MS) metabolomics data is discussed. Feature selection methods utilized for multivariate data analysis are briefly reviewed, including feature clustering and computation of latent vectors using spectral methods. We propose that the time-course of metabolite changes in samples stored at different temperatures may be utilized to identify changing-metabolite-to-stable-metabolite ratios as markers of sample fitness. Tolerance intervals may be computed to characterize these ratios among fresh samples. In order to discover additional structure in the data relevant to sample fitness, we propose using data labeled according to these ratios to train a Dirichlet process mixture model (DPMM) for assessing sample fitness. DPMMs are highly intuitive since they model the metabolite levels in a sample as arising from a combination of processes including, e.g., normal biological processes and degradation- or contamination-inducing processes. The outputs of a DPMM are probabilities that a sample is associated with a given process, and these probabilities may be incorporated into a final classifier for sample fitness.

📄 PDF Abstract BibTeX arXiv:1506.04443

Code (0)

등록된 구현이 없습니다.

Tasks

feature selection

Similar Papers 제목 키워드 기반

MS2MetGAN: Latent-space adversarial training for metabolite-spectrum matching in MS/MS database search

2026-03-07 · Meng Tsai, Alexzander Dwyer, Estelle Nuckels, Yingfeng Wang arxiv

Database search is a widely used approach for identifying metabolites from tandem mass spectra (MS/MS). In this strategy, an experimental spectrum is matched against a user-specified database of candidate metabolites, an…

Using statistical techniques and replication samples for imputation of metabolite missing values

2019-05-12

Background: Data preparation, such as missing values imputation and transformation, is the first step in any data analysis and requires crucial attention. Particularly, analysis of metabolites demands more preparation si…

ImputationMissing Values

Pathway-Activity Likelihood Analysis and Metabolite Annotation for Untargeted Metabolomics using Probabilistic Modeling

2019-12-12 · Ramtin Hosseini, Neda Hassanpour, Li-Ping Liu, Soha Hassoun

Motivation: Untargeted metabolomics comprehensively characterizes small molecules and elucidates activities of biochemical pathways within a biological sample. Despite computational advances, interpreting collected measu…

Predicting pathways for old and new metabolites through clustering

2022-11-28 · Thiru Siddharth, Nathan Lewis

The diverse metabolic pathways are fundamental to all living organisms, as they harvest energy, synthesize biomass components, produce molecules to interact with the microenvironment, and neutralize toxins. While discove…

Clustering

Metabolomic profiles in Jamaican children with and without autism spectrum disorder

2024-03-11 · Akram Yazdani, Maureen Samms-Vaughan, Sepideh Saroukhani, Jan Bressler 외

Autism spectrum disorder (ASD) is a complex neurodevelopmental condition with a wide range of behavioral and cognitive impairments. While genetic and environmental factors are known to contribute to its etiology, the und…

ImputationMissing Values