Unknown

Dataset Information

0

MsDetector: toward a standard computational tool for DNA microsatellites detection.


ABSTRACT: Microsatellites (MSs) are DNA regions consisting of repeated short motif(s). MSs are linked to several diseases and have important biomedical applications. Thus, researchers have developed several computational tools to detect MSs. However, the currently available tools require adjusting many parameters, or depend on a list of motifs or on a library of known MSs. Therefore, two laboratories analyzing the same sequence with the same computational tool may obtain different results due to the user-adjustable parameters. Recent studies have indicated the need for a standard computational tool for detecting MSs. To this end, we applied machine-learning algorithms to develop a tool called MsDetector. The system is based on a hidden Markov model and a general linear model. The user is not obligated to optimize the parameters of MsDetector. Neither a list of motifs nor a library of known MSs is required. MsDetector is memory- and time-efficient. We applied MsDetector to several species. MsDetector located the majority of MSs found by other widely used tools. In addition, MsDetector identified novel MSs. Furthermore, the system has a very low false-positive rate resulting in a precision of up to 99%. MsDetector is expected to produce consistent results across studies analyzing the same sequence.

SUBMITTER: Girgis HZ 

PROVIDER: S-EPMC3592430 | biostudies-literature | 2013 Jan

REPOSITORIES: biostudies-literature

altmetric image

Publications

MsDetector: toward a standard computational tool for DNA microsatellites detection.

Girgis Hani Z HZ   Sheetlin Sergey L SL  

Nucleic acids research 20121002 1


Microsatellites (MSs) are DNA regions consisting of repeated short motif(s). MSs are linked to several diseases and have important biomedical applications. Thus, researchers have developed several computational tools to detect MSs. However, the currently available tools require adjusting many parameters, or depend on a list of motifs or on a library of known MSs. Therefore, two laboratories analyzing the same sequence with the same computational tool may obtain different results due to the user-  ...[more]

Similar Datasets

| S-EPMC5858232 | biostudies-literature
| S-EPMC3222720 | biostudies-literature
| S-EPMC2916834 | biostudies-literature
| S-EPMC6506443 | biostudies-literature
| S-EPMC6551572 | biostudies-literature
| S-EPMC7072524 | biostudies-literature
| S-EPMC7279028 | biostudies-literature
| S-EPMC4762524 | biostudies-literature
| S-EPMC194918 | biostudies-literature
| S-EPMC7115099 | biostudies-literature