paper-with-me

홈 › Papers

Distinguishing Enzyme Structures from Non-enzymes Without Alignments

2003-07-18 · Journal of Molecular Biology 2003 7 · Paul D.Dobson, Andrew J.Doig

The ability to predict protein function from structure is becoming increasingly important as the number of structures resolved is growing more rapidly than our capacity to study function. Current methods for predicting protein function are mostly reliant on identifying a similar protein of known function. For proteins that are highly dissimilar or are only similar to proteins also lacking functional annotations, these methods fail. Here, we show that protein function can be predicted as enzymatic or not without resorting to alignments. We describe 1178 high-resolution proteins in a structurally non-redundant subset of the Protein Data Bank using simple features such as secondary-structure content, amino acid propensities, surface properties and ligands. The subset is split into two functional groupings, enzymes and non-enzymes. We use the support vector machine-learning algorithm to develop models that are capable of assigning the protein class. Validation of the method shows that the function can be predicted to an accuracy of 77% using 52 features to describe each protein. An adaptive search of possible subsets of features produces a simplified model based on 36 features that predicts at an accuracy of 80%. We compare the method to sequence-based methods that also avoid calculating alignments and predict a recently released set of unrelated proteins. The most useful features for distinguishing enzymes from non-enzymes are secondary-structure content, amino acid frequencies, number of disulphide bonds and size of the largest cleft. This method is applicable to any structure as it does not require the identification of sequence or structural similarity to a protein of known function.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Graph Classification

Similar Papers 제목 키워드 기반

Identification of functionally related enzymes by learning-to-rank methods

2014-05-17 · Michiel Stock, Thomas Fober, Eyke Hüllermeier, Serghei Glinca 외

Enzyme sequences and structures are routinely used in the biological sciences as queries to search for functionally related enzymes in online databases. To this end, one usually departs from some notion of similarity, co…

Learning-To-Rank

Structures, Functions, and Mechanisms of Filament Forming Enzymes: A Renaissance of Enzyme Filamentation

2019-09-28

Filament formation by non-cytoskeletal enzymes has been known for decades, yet only relatively recently has its wide-spread role in enzyme regulation and biology come to be appreciated. This comprehensive review summariz…

UniZyme: A Unified Protein Cleavage Site Predictor Enhanced with Enzyme Active-Site Knowledge

2025-02-10 · Chenao Li, Shuo Yan, Enyan Dai

Enzyme-catalyzed protein cleavage is essential for many biological functions. Accurate prediction of cleavage sites can facilitate various applications such as drug development, enzyme design, and a deeper understanding …

General Multimodal Protein Design Enables DNA-Encoding of Chemistry

2026-04-06 · Jarrid Rector-Brooks, Théophile Lambert, Marta Skreta, Daniel Roth 외 arxiv

Evolution is an extraordinary engine for enzymatic diversity, yet the chemistry it has explored remains a narrow slice of what DNA can encode. Deep generative models can design new proteins that bind ligands, but none ha…

Protein Design

An Investigation of Hepatitis B Virus Genome using Markov Models

2023-11-12 · Khadijeh, Jahanian, Elnaz Shalbafian, Morteza Saberi 외

The human genome encodes a family of editing enzymes known as APOBEC3 (apolipoprotein B mRNA editing enzyme, catalytic polypeptide-like 3). Several family members, such as APO-BEC3G, APOBEC3F, and APOBEC3H haplotype II, …