Classification of Protein Crystallization X-Ray Images Using Major Convolutional Neural Network Architectures
The generation of protein crystals is necessary for the study of protein molecular function and structure. This is done empirically by processing large numbers of crystallization trials and inspecting them regularly in search of those with forming crystals. To avoid missing the hard-gained crystals, this visual inspection of the trial X-ray images is done manually as opposed to the existing less accurate machine learning methods. To achieve higher accuracy for automation, we applied some of the most successful convolutional neural networks (ResNet, Inception, VGG, and AlexNet) for 10-way classification of the X-ray images. We showed that substantial classification accuracy is gained by using such networks compared to two simpler ones previously proposed for this purpose. The best accuracy was obtained from ResNet (81.43%), which corresponds to a missed crystal rate of 5.9%. This rate could be lowered to less than 0.1% by using a top-3 classification strategy. Our dataset consisted of 486,000 internally annotated images, which was augmented to more than a million to address class imbalance. We also provide a label-wise analysis of the results, identifying the main sources of error and inaccuracy.
Code (0)
등록된 구현이 없습니다.
Tasks
General ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Classification of X-Ray Protein Crystallization Using Deep Convolutional Neural Networks with a Finder Module
Recently, deep convolutional neural networks have shown good results for image recognition. In this paper, we use convolutional neural networks with a finder module, which discovers the important region for recognition a…
General ClassificationClassification of crystallization outcomes using deep convolutional neural networks
The Machine Recognition of Crystallization Outcomes (MARCO) initiative has assembled roughly half a million annotated images of macromolecular crystallization experiments from various sources and setups. Here, state-of-t…
BIG-bench Machine LearningClassificationGeneral ClassificationCrystallization via tubing microfluidics permits both in situ and ex situ X-ray diffraction
We used a microfluidic platform to address the problems of obtaining diffraction quality crystals and crystal handling during transfer to the X-ray diffractometer. We optimize crystallization conditions of a pharmaceutic…
Better results with homogeneous biological macromolecules
Pure and homogeneous biological macromolecules (i.e. proteins, nucleic acids, protein-protein or protein-nucleic acid complexes, and functional assemblies such as ribosomes and viruses) are the key for consistent and rel…
Learning multi-scale functional representations of proteins from single-cell microscopy data
Protein function is inherently linked to its localization within the cell, and fluorescent microscopy data is an indispensable resource for learning representations of proteins. Despite major developments in molecular re…
molecular representationRepresentation Learning