paper-with-me

홈 › Papers

Developing a data analysis pipeline for automated protein profiling in immunology

2022-01-16 · Dmytro Fishman

Accurate information about protein content in the organism is instrumental for a better understanding of human biology and disease mechanisms. While the presence of certain types of proteins can be life-threatening, the abundance of others is an essential condition for an individual's overall well-being. Protein microarray is a technology that enables the quantification of thousands of proteins in hundreds of human samples in a parallel manner. In a series of studies involving protein microarrays, we have explored and implemented various data science methods for all-around analysing of these data. This analysis has enabled the identification and characterisation of proteins targeted by the autoimmune reaction in patients with the APS1 condition. We have also assessed the utility of applying machine learning methods alongside statistical tests in a study based on protein expression data to evaluate potential biomarkers for endometriosis. The keystone of this work is a web-tool PAWER. PAWER implements relevant computational methods, and provides a semi-automatic way to run the analysis of protein microarray data online in a drag-and-drop and click-and-play style. The source code of the tool is publicly available. The work that laid the foundation of this thesis has been instrumental for a number of subsequent studies of human disease and also inspired a contribution to refining standards for validation of machine learning methods in biology.

📄 PDF Abstract BibTeX arXiv:2201.06074

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards fully automated protein structure elucidation with NMR spectroscopy

2018-07-31 · Piotr Klukowski, Adam Gonczarek

Nuclear magnetic resonance (NMR) spectroscopy is one of the leading techniques for protein studies. The method features a number of properties, allowing to explain macromolecular interactions mechanistically and resolve …

Combinatorial Optimization

Navigating Eukaryotic Genome Annotation Pipelines: A Route Map to BRAKER, Galba, and TSEBRA

2024-03-28 · Tomáš Brůna, Lars Gabriel, Katharina J. Hoff

Annotating the structure of protein-coding genes represents a major challenge in the analysis of eukaryotic genomes. This task sets the groundwork for subsequent genomic studies aimed at understanding the functions of in…

ProteinAE: Protein Diffusion Autoencoders for Structure Encoding

2025-10-12 · Shaoning Li, Le Zhuo, Yusong Wang, Mingyu Li 외 arxiv

Developing effective representations of protein structures is essential for advancing protein science, particularly for protein generative modeling. Current approaches often grapple with the complexities of the SE(3) man…

Automating MD simulations for Proteins using Large language Models: NAMD-Agent

2025-07-10 · Achuth Chandrasekhar, Amir Barati Farimani

Molecular dynamics simulations are an essential tool in understanding protein structure, dynamics, and function at the atomic level. However, preparing high quality input files for MD simulations can be a time consuming …

Code GenerationNavigate

Automated Machine Learning Pipeline: Large Language Models-Assisted Automated Dataset Generation for Training Machine-Learned Interatomic Potentials

2025-09-25 · Adam Lahouari, Jutta Rogal, Mark E. Tuckerman arxiv

Machine learning interatomic potentials (MLIPs) have become powerful tools to extend molecular simulations beyond the limits of quantum methods, offering near-quantum accuracy at much lower computational cost. Yet, devel…