Home LiteratureArticle Details
PMID: 12682369 Published · ppublish English Journal Article Research Support, Non-U.S. Gov't

GenDB--an open source genome annotation system for prokaryote genomes.

Nucleic acids research ·Vol. 31 ·No. 8 ·2003-04-15 ·Pages 2187-95

Meyer F, Goesmann A, McHardy AC, Bartels D, Bekel T, Clausen J, Kalinowski J, Linke B, Rupp O, Giegerich R, Pühler A

Abstract

The flood of sequence data resulting from the large number of current genome projects has increased the need for a flexible, open source genome annotation system, which so far has not existed. To account for the individual needs of different projects, such a system should be modular and easily extensible. We present a genome annotation system for prokaryote genomes, which is well tested and readily adaptable to different tasks. The modular system was developed using an object-oriented approach, and it relies on a relational database backend. Using a well defined application programmers interface (API), the system can be linked easily to other systems. GenDB supports manual as well as automatic annotation strategies. The software currently is in use in more than a dozen microbial genome annotation projects. In addition to its use as a production genome annotation system, it can be employed as a flexible framework for the large-scale evaluation of different annotation strategies. The system is open source.

MeSH Terms
Amino Acid Sequence Bacteria/genetics Base Sequence Computational Biology/methods Genome Genome, Bacterial Internet Molecular Sequence Data Prokaryotic Cells/metabolism Sequence Homology, Amino Acid Software
Authors & Affiliations
11 authors, click to expand affiliations / ORCID
Meyer Folker
Center for Genome Research, Department of Biology, Bielefeld University, Bielefeld, Germany. fm@genetik.uni-bielefeld.de
Goesmann Alexander
McHardy Alice C
Bartels Daniela
Bekel Thomas
Clausen Jörn
Kalinowski Jörn
Linke Burkhard
Rupp Oliver
Giegerich Robert
Pühler Alfred
References (19)
19 references, click to expand
  1. The InterPro database, an integrated documentation resource for protein families, domains and functional sites.
    Nucleic Acids Res. 2001 Jan 1;29(1):37-40 PMID: 11125043
  2. Artemis: sequence visualization and annotation.
    Bioinformatics. 2000 Oct;16(10):944-5 PMID: 11120685
  3. The composite genome of the legume symbiont Sinorhizobium meliloti.
    Science. 2001 Jul 27;293(5530):668-72 PMID: 11474104
  4. PathFinder: reconstruction and dynamic visualization of metabolic pathways.
    Bioinformatics. 2002 Jan;18(1):124-9 PMID: 11836220
  5. The complete nucleotide sequence and environmental distribution of the cryptic, conjugative, broad-host-range plasmid pIPO2 isolated from bacteria of the wheat rhizosphere.
    Microbiology. 2002 Jun;148(Pt 6):1637-53 PMID: 12055285
  6. The 79,370-bp conjugative plasmid pB4 consists of an IncP-1beta backbone loaded with a chromate resistance transposon, the strA-strB streptomycin resistance gene pair, the oxacillinase gene bla(NPS-1), and a tripartite antibiotic efflux system of the resistance-nodulation-division family.
    Mol Genet Genomics. 2003 Feb;268(5):570-84 PMID: 12589432
  7. Automated assembly of protein blocks for database searching.
    Nucleic Acids Res. 1991 Dec 11;19(23):6565-72 PMID: 1754394
  8. SRS--an indexing and retrieval tool for flat file data libraries.
    Comput Appl Biosci. 1993 Feb;9(1):49-57 PMID: 8435768
  9. MAGPIE: automated genome interpretation.
    Trends Genet. 1996 Feb;12(2):76-8 PMID: 8851977
  10. tRNAscan-SE: a program for improved detection of transfer RNA genes in genomic sequence.
    Nucleic Acids Res. 1997 Mar 1;25(5):955-64 PMID: 9023104
  11. Identification of prokaryotic and eukaryotic signal peptides and prediction of their cleavage sites.
    Protein Eng. 1997 Jan;10(1):1-6 PMID: 9051728
  12. Gapped BLAST and PSI-BLAST: a new generation of protein database search programs.
    Nucleic Acids Res. 1997 Sep 1;25(17):3389-402 PMID: 9254694
  13. Profile hidden Markov models.
    Bioinformatics. 1998;14(9):755-63 PMID: 9918945
  14. The use of gene clusters to infer functional coupling.
    Proc Natl Acad Sci U S A. 1999 Mar 16;96(6):2896-901 PMID: 10077608
  15. CRITICA: coding region identification tool invoking comparative analysis.
    Mol Biol Evol. 1999 Apr;16(4):512-24 PMID: 10331277
  16. Automated genome sequence analysis and annotation.
    Bioinformatics. 1999 May;15(5):391-412 PMID: 10366660
  17. Improved microbial gene identification with GLIMMER.
    Nucleic Acids Res. 1999 Dec 1;27(23):4636-41 PMID: 10556321
  18. WIT: integrated system for high-throughput genome sequence analysis and metabolic reconstruction.
    Nucleic Acids Res. 2000 Jan 1;28(1):123-5 PMID: 10592199
  19. Functional and structural genomics using PEDANT.
    Bioinformatics. 2001 Jan;17(1):44-57 PMID: 11222261
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
1362-4962
Published
2003-04-15
Pages
2187-95
Language
English
Region
England
NLM ID
0411011
PMCID
PMC153740
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com