paper-with-me

홈 › Papers

ProtRank: Bypassing the imputation of missing values in differential expression analysis of proteomic data

2019-09-30

Data from discovery proteomic and phosphoproteomic experiments typically include missing values that correspond to proteins that have not been identified in the analyzed sample. Replacing the missing values with random numbers, a process known as "imputation", avoids apparent infinite fold-change values. However, the procedure comes at a cost: Imputing a large number of missing values has the potential to significantly impact the results of the subsequent differential expression analysis. We propose a method that identifies differentially expressed proteins by ranking their observed changes with respect to the changes observed for other proteins. Missing values are taken into account by this method directly, without the need to impute them. We illustrate the performance of the new method on two distinct datasets and show that it is robust to missing values and, at the same time, provides results that are otherwise similar to those obtained with edgeR which is a state-of-art differential expression analysis method. The new method for the differential expression analysis of proteomic data is available as an easy to use Python package.

📄 PDF Abstract BibTeX arXiv:1909.13667

Code (1)

8medom/ProtRank 공식 구현

Tasks

ImputationMissing Values

Similar Papers 제목 키워드 기반

Quantum-Inspired Optimization Process for Data Imputation

2025-05-07 · Nishikanta Mohanty, Bikash K. Behera, Badshah Mukherjee, Christopher Ferrie

Data imputation is a critical step in data pre-processing, particularly for datasets with missing or unreliable values. This study introduces a novel quantum-inspired imputation framework evaluated on the UCI Diabetes da…

ImputationMissing Values

RDIS: Random Drop Imputation with Self-Training for Incomplete Time Series Data

2020-10-20 · Tae-Min Choi, Ji-Su Kang, Jong-Hwan Kim

Time-series data with missing values are commonly encountered in many fields, such as healthcare, meteorology, and robotics. The imputation aims to fill the missing values with valid values. Most imputation methods train…

ImputationMissing ValuesTime SeriesTime Series Analysis+1

Impute-MACFM: Imputation based on Mask-Aware Flow Matching

2025-09-27 · Dengyi Liu, Honggang Wang, Hua Fang arxiv

Tabular data are central to many applications, especially longitudinal data in healthcare, where missing values are common, undermining model fidelity and reliability. Prior imputation methods either impose restrictive a…

Revisiting the thorny issue of missing values in single-cell proteomics

2023-04-13 · Christophe Vanderaa, Laurent Gatto

Missing values are a notable challenge when analysing mass spectrometry-based proteomics data. While the field is still actively debating on the best practices, the challenge increased with the emergence of mass spectrom…

ImputationManagementMissing Values

Using statistical techniques and replication samples for imputation of metabolite missing values

2019-05-12

Background: Data preparation, such as missing values imputation and transformation, is the first step in any data analysis and requires crucial attention. Particularly, analysis of metabolites demands more preparation si…

ImputationMissing Values