Unknown

Dataset Information

0

Enhancing the functional content of eukaryotic protein interaction networks.


ABSTRACT: Protein interaction networks are a promising type of data for studying complex biological systems. However, despite the rich information embedded in these networks, these networks face important data quality challenges of noise and incompleteness that adversely affect the results obtained from their analysis. Here, we apply a robust measure of local network structure called common neighborhood similarity (CNS) to address these challenges. Although several CNS measures have been proposed in the literature, an understanding of their relative efficacies for the analysis of interaction networks has been lacking. We follow the framework of graph transformation to convert the given interaction network into a transformed network corresponding to a variety of CNS measures evaluated. The effectiveness of each measure is then estimated by comparing the quality of protein function predictions obtained from its corresponding transformed network with those from the original network. Using a large set of human and fly protein interactions, and a set of over 100 GO terms for both, we find that several of the transformed networks produce more accurate predictions than those obtained from the original network. In particular, the HC.cont measure and other continuous CNS measures perform well this task, especially for large networks. Further investigation reveals that the two major factors contributing to this improvement are the abilities of CNS measures to prune out noisy edges and enhance functional coherence in the transformed networks.

SUBMITTER: Pandey G 

PROVIDER: S-EPMC4183583 | biostudies-literature | 2014

REPOSITORIES: biostudies-literature

altmetric image

Publications

Enhancing the functional content of eukaryotic protein interaction networks.

Pandey Gaurav G   Arora Sonali S   Manocha Sahil S   Whalen Sean S  

PloS one 20141002 10


Protein interaction networks are a promising type of data for studying complex biological systems. However, despite the rich information embedded in these networks, these networks face important data quality challenges of noise and incompleteness that adversely affect the results obtained from their analysis. Here, we apply a robust measure of local network structure called common neighborhood similarity (CNS) to address these challenges. Although several CNS measures have been proposed in the l  ...[more]

Similar Datasets

| S-EPMC1797819 | biostudies-literature
| S-EPMC3924044 | biostudies-literature
| S-EPMC3524085 | biostudies-other
| S-EPMC3467740 | biostudies-literature
| S-EPMC5701033 | biostudies-literature
| S-EPMC3459874 | biostudies-literature
| S-EPMC3323588 | biostudies-literature
| S-EPMC4762622 | biostudies-literature
| S-EPMC4606133 | biostudies-literature
| S-EPMC3079720 | biostudies-literature