Unknown

Dataset Information

0

KEGG as a reference resource for gene and protein annotation.


ABSTRACT: KEGG (http://www.kegg.jp/ or http://www.genome.jp/kegg/) is an integrated database resource for biological interpretation of genome sequences and other high-throughput data. Molecular functions of genes and proteins are associated with ortholog groups and stored in the KEGG Orthology (KO) database. The KEGG pathway maps, BRITE hierarchies and KEGG modules are developed as networks of KO nodes, representing high-level functions of the cell and the organism. Currently, more than 4000 complete genomes are annotated with KOs in the KEGG GENES database, which can be used as a reference data set for KO assignment and subsequent reconstruction of KEGG pathways and other molecular networks. As an annotation resource, the following improvements have been made. First, each KO record is re-examined and associated with protein sequence data used in experiments of functional characterization. Second, the GENES database now includes viruses, plasmids, and the addendum category for functionally characterized proteins that are not represented in complete genomes. Third, new automatic annotation servers, BlastKOALA and GhostKOALA, are made available utilizing the non-redundant pangenome data set generated from the GENES database. As a resource for translational bioinformatics, various data sets are created for antimicrobial resistance and drug interaction networks.

SUBMITTER: Kanehisa M 

PROVIDER: S-EPMC4702792 | biostudies-literature | 2016 Jan

REPOSITORIES: biostudies-literature

altmetric image

Publications

KEGG as a reference resource for gene and protein annotation.

Kanehisa Minoru M   Sato Yoko Y   Kawashima Masayuki M   Furumichi Miho M   Tanabe Mao M  

Nucleic acids research 20151017 D1


KEGG (http://www.kegg.jp/ or http://www.genome.jp/kegg/) is an integrated database resource for biological interpretation of genome sequences and other high-throughput data. Molecular functions of genes and proteins are associated with ortholog groups and stored in the KEGG Orthology (KO) database. The KEGG pathway maps, BRITE hierarchies and KEGG modules are developed as networks of KO nodes, representing high-level functions of the cell and the organism. Currently, more than 4000 complete geno  ...[more]

Similar Datasets

| S-EPMC308797 | biostudies-literature
| S-EPMC2324097 | biostudies-literature
| S-EPMC2894510 | biostudies-literature
| S-EPMC2703891 | biostudies-literature
| S-EPMC2686502 | biostudies-literature
| S-EPMC6668538 | biostudies-literature
| S-EPMC8976094 | biostudies-literature
| S-EPMC2686469 | biostudies-literature
| S-EPMC3245047 | biostudies-literature
| S-EPMC4602055 | biostudies-literature