Unknown

Dataset Information

0

An ANN-GA model based promoter prediction in Arabidopsis thaliana using tilling microarray data.


ABSTRACT: Identification of promoter region is an important part of gene annotation. Identification of promoters in eukaryotes is important as promoters modulate various metabolic functions and cellular stress responses. In this work, a novel approach utilizing intensity values of tilling microarray data for a model eukaryotic plant Arabidopsis thaliana, was used to specify promoter region from non-promoter region. A feed-forward back propagation neural network model supported by genetic algorithm was employed to predict the class of data with a window size of 41. A dataset comprising of 2992 data vectors representing both promoter and non-promoter regions, chosen randomly from probe intensity vectors for whole genome of Arabidopsis thaliana generated through tilling microarray technique was used. The classifier model shows prediction accuracy of 69.73% and 65.36% on training and validation sets, respectively. Further, a concept of distance based class membership was used to validate reliability of classifier, which showed promising results. The study shows the usability of micro-array probe intensities to predict the promoter regions in eukaryotic genomes.

SUBMITTER: Mishra H 

PROVIDER: S-EPMC3159145 | biostudies-literature | 2011

REPOSITORIES: biostudies-literature

altmetric image

Publications

An ANN-GA model based promoter prediction in Arabidopsis thaliana using tilling microarray data.

Mishra Hrishikesh H   Singh Nitya N   Misra Krishna K   Lahiri Tapobrata T  

Bioinformation 20110606 6


Identification of promoter region is an important part of gene annotation. Identification of promoters in eukaryotes is important as promoters modulate various metabolic functions and cellular stress responses. In this work, a novel approach utilizing intensity values of tilling microarray data for a model eukaryotic plant Arabidopsis thaliana, was used to specify promoter region from non-promoter region. A feed-forward back propagation neural network model supported by genetic algorithm was emp  ...[more]

Similar Datasets

| S-EPMC4596684 | biostudies-literature
| S-EPMC3124792 | biostudies-literature
2007-03-24 | GSE7353 | GEO
| S-EPMC1794575 | biostudies-literature
| S-EPMC7240361 | biostudies-literature
| S-EPMC2709567 | biostudies-literature