Unknown

Dataset Information

0

Statistical evaluation of improvement in RNA secondary structure prediction.


ABSTRACT: With discovery of diverse roles for RNA, its centrality in cellular functions has become increasingly apparent. A number of algorithms have been developed to predict RNA secondary structure. Their performance has been benchmarked by comparing structure predictions to reference secondary structures. Generally, algorithms are compared against each other and one is selected as best without statistical testing to determine whether the improvement is significant. In this work, it is demonstrated that the prediction accuracies of methods correlate with each other over sets of sequences. One possible reason for this correlation is that many algorithms use the same underlying principles. A set of benchmarks published previously for programs that predict a structure common to three or more sequences is statistically analyzed as an example to show that it can be rigorously evaluated using paired two-sample t-tests. Finally, a pipeline of statistical analyses is proposed to guide the choice of data set size and performance assessment for benchmarks of structure prediction. The pipeline is applied using 5S rRNA sequences as an example.

SUBMITTER: Xu Z 

PROVIDER: S-EPMC3287165 | biostudies-literature | 2012 Feb

REPOSITORIES: biostudies-literature

altmetric image

Publications

Statistical evaluation of improvement in RNA secondary structure prediction.

Xu Zhenjiang Z   Almudevar Anthony A   Mathews David H DH  

Nucleic acids research 20111201 4


With discovery of diverse roles for RNA, its centrality in cellular functions has become increasingly apparent. A number of algorithms have been developed to predict RNA secondary structure. Their performance has been benchmarked by comparing structure predictions to reference secondary structures. Generally, algorithms are compared against each other and one is selected as best without statistical testing to determine whether the improvement is significant. In this work, it is demonstrated that  ...[more]

Similar Datasets

| S-EPMC1383571 | biostudies-literature
| S-EPMC297010 | biostudies-literature
| S-EPMC3629938 | biostudies-literature
| S-EPMC3819574 | biostudies-literature
| S-EPMC2874162 | biostudies-literature
| S-EPMC2536673 | biostudies-literature
| S-EPMC8318245 | biostudies-literature
| S-EPMC3667108 | biostudies-literature
| S-EPMC6439966 | biostudies-literature
| S-EPMC29728 | biostudies-literature