Unknown

Dataset Information

0

Network-based protein structural classification.


ABSTRACT: Experimental determination of protein function is resource-consuming. As an alternative, computational prediction of protein function has received attention. In this context, protein structural classification (PSC) can help, by allowing for determining structural classes of currently unclassified proteins based on their features, and then relying on the fact that proteins with similar structures have similar functions. Existing PSC approaches rely on sequence-based or direct three-dimensional (3D) structure-based protein features. By contrast, we first model 3D structures of proteins as protein structure networks (PSNs). Then, we use network-based features for PSC. We propose the use of graphlets, state-of-the-art features in many research areas of network science, in the task of PSC. Moreover, because graphlets can deal only with unweighted PSNs, and because accounting for edge weights when constructing PSNs could improve PSC accuracy, we also propose a deep learning framework that automatically learns network features from weighted PSNs. When evaluated on a large set of approximately 9400 CATH and approximately 12 800 SCOP protein domains (spanning 36 PSN sets), the best of our proposed approaches are superior to existing PSC approaches in terms of accuracy, with comparable running times. Our data and code are available at https://doi.org/10.5281/zenodo.3787922.

SUBMITTER: Newaz K 

PROVIDER: S-EPMC7353965 | biostudies-literature | 2020 Jun

REPOSITORIES: biostudies-literature

altmetric image

Publications

Network-based protein structural classification.

Newaz Khalique K   Ghalehnovi Mahboobeh M   Rahnama Arash A   Antsaklis Panos J PJ   Milenković Tijana T  

Royal Society open science 20200603 6


Experimental determination of protein function is resource-consuming. As an alternative, computational prediction of protein function has received attention. In this context, protein structural classification (PSC) can help, by allowing for determining structural classes of currently unclassified proteins based on their features, and then relying on the fact that proteins with similar structures have similar functions. Existing PSC approaches rely on sequence-based or direct three-dimensional (3  ...[more]

Similar Datasets

| S-EPMC5370107 | biostudies-literature
| S-EPMC7916854 | biostudies-literature
| S-EPMC1347461 | biostudies-literature
| S-EPMC2063581 | biostudies-literature
2023-12-02 | GSE248664 | GEO
| S-EPMC3093400 | biostudies-literature
| S-EPMC3268291 | biostudies-literature
| S-EPMC2754988 | biostudies-literature
| S-EPMC1636673 | biostudies-literature
| S-EPMC2873827 | biostudies-literature