paper-with-me

홈 › Papers

Weighted Scaling Approach for Metabolomics Data Analysis

2022-08-01 · Biplab Biswas, Nishith Kumar, Md Aminul Hoque, Md ashad Alam

Systematic variation is a common issue in metabolomics data analysis. Therefore, different scaling and normalization techniques are used to preprocess the data for metabolomics data analysis. Although several scaling methods are available in the literature, however, choice of scaling, transformation and/or normalization technique influence the further statistical analysis. It is challenging to choose the appropriate scaling technique for downstream analysis to get accurate results or to make a proper decision. Moreover, the existing scaling techniques are sensitive to outliers or extreme values. To fill the gap, our objective is to introduce a robust scaling approach that is not influenced by outliers as well as provides more accurate results for downstream analysis. Here, we introduced a new weighted scaling approach that is robust against outliers however, where no additional outlier detection/treatment step is needed in data preprocessing and also compared it with the conventional scaling and normalization techniques through artificial and real metabolomics datasets. We evaluated the performance of the proposed method in comparison to the other existing conventional scaling techniques using metabolomics data analysis in both the absence and presence of different percentages of outliers. Results show that in most cases, the proposed scaling technique performs better than the traditional scaling methods in both the absence and presence of outliers. The proposed method improves the further downstream metabolomics analysis. The R function of the proposed robust scaling method is available at https://github.com/nishithkumarpaul/robustScaling/blob/main/wscaling.R

📄 PDF Abstract BibTeX arXiv:2208.00603

Code (1)

nishithkumarpaul/robustscaling 공식 구현

Tasks

Outlier Detection

Similar Papers 제목 키워드 기반

Dynamical models for metabolomics data integration

2021-05-21 · Polina Lakrisenko, Daniel Weindl

As metabolomics datasets are becoming larger and more complex, there is an increasing need for model-based data integration and analysis to optimally leverage these data. Dynamical models of metabolism allow for the inte…

Data Integration

Bioinformatics Analysis of Metabolomics Data Unveils Association of Metabolic Signatures with Methylation in Breast Cancer

2019-12-29 · Fadhl M. Alakwaa, Lana X Garmire, Masha G. Savelieff

Breast cancer (BC) contributes the highest global cancer mortality in women. BC tumors are highly heterogeneous, so subtyping by cell-surface markers is inadequate. Omics-driven tumor stratification is urgently needed to…

Classifying Dry Eye Disease Patients from Healthy Controls Using Machine Learning and Metabolomics Data

2024-06-20 · Sajad Amouei Sheshkal, Morten Gundersen, Michael Alexander Riegler, Øygunn Aass Utheim 외

Dry eye disease is a common disorder of the ocular surface, leading patients to seek eye care. Clinical signs and symptoms are currently used to diagnose dry eye disease. Metabolomics, a method for analyzing biological s…

regressionSpecificity

MetaBench: A Multi-task Benchmark for Assessing LLMs in Metabolomics

2025-10-16 · Yuxing Lu, Xukai Zhao, J. Ben Tamo, Micky C. Nnamdi 외 arxiv

Large Language Models (LLMs) have demonstrated remarkable capabilities on general text; however, their proficiency in specialized scientific domains that require deep, interconnected knowledge remains largely uncharacter…

Text Generation

Multi-View Variational Autoencoder for Missing Value Imputation in Untargeted Metabolomics

2023-10-12 · Chen Zhao, Kuan-Jui Su, Chong Wu, Xuewei Cao 외

Background: Missing data is a common challenge in mass spectrometry-based metabolomics, which can lead to biased and incomplete analyses. The integration of whole-genome sequencing (WGS) data with metabolomics data has e…

Data IntegrationImputationMissing Values