Unknown

Dataset Information

0

Discovering protein-DNA binding sequence patterns using association rule mining.


ABSTRACT: Protein-DNA bindings between transcription factors (TFs) and transcription factor binding sites (TFBSs) play an essential role in transcriptional regulation. Over the past decades, significant efforts have been made to study the principles for protein-DNA bindings. However, it is considered that there are no simple one-to-one rules between amino acids and nucleotides. Many methods impose complicated features beyond sequence patterns. Protein-DNA bindings are formed from associated amino acid and nucleotide sequence pairs, which determine many functional characteristics. Therefore, it is desirable to investigate associated sequence patterns between TFs and TFBSs. With increasing computational power, availability of massive experimental databases on DNA and proteins, and mature data mining techniques, we propose a framework to discover associated TF-TFBS binding sequence patterns in the most explicit and interpretable form from TRANSFAC. The framework is based on association rule mining with Apriori algorithm. The patterns found are evaluated by quantitative measurements at several levels on TRANSFAC. With further independent verifications from literatures, Protein Data Bank and homology modeling, there are strong evidences that the patterns discovered reveal real TF-TFBS bindings across different TFs and TFBSs, which can drive for further knowledge to better understand TF-TFBS bindings.

SUBMITTER: Leung KS 

PROVIDER: S-EPMC2965231 | biostudies-literature | 2010 Oct

REPOSITORIES: biostudies-literature

altmetric image

Publications

Discovering protein-DNA binding sequence patterns using association rule mining.

Leung Kwong-Sak KS   Wong Ka-Chun KC   Chan Tak-Ming TM   Wong Man-Hon MH   Lee Kin-Hong KH   Lau Chi-Kong CK   Tsui Stephen K W SK  

Nucleic acids research 20100606 19


Protein-DNA bindings between transcription factors (TFs) and transcription factor binding sites (TFBSs) play an essential role in transcriptional regulation. Over the past decades, significant efforts have been made to study the principles for protein-DNA bindings. However, it is considered that there are no simple one-to-one rules between amino acids and nucleotides. Many methods impose complicated features beyond sequence patterns. Protein-DNA bindings are formed from associated amino acid and  ...[more]

Similar Datasets

| S-EPMC4059059 | biostudies-literature
| S-EPMC4530207 | biostudies-literature
| S-EPMC3064845 | biostudies-other
| S-EPMC5249029 | biostudies-literature
| S-EPMC4849775 | biostudies-literature
| S-EPMC6473086 | biostudies-literature
| S-EPMC5291444 | biostudies-literature
| S-EPMC6866600 | biostudies-literature
| S-EPMC2718668 | biostudies-literature
| S-EPMC7183129 | biostudies-literature