Home LiteratureArticle Details
PMID: 22123736 Published · ppublish English Journal Article Research Support, N.I.H., Extramural Research Support, Non-U.S. Gov't

The UniProt-GO Annotation database in 2011.

Nucleic acids research ·Vol. 40 ·No. Database issue ·2012-01-00 ·Pages D565-70

Dimmer EC, Huntley RP, Alam-Faruque Y, Sawford T, O'Donovan C, Martin MJ, Bely B, Browne P, Mun Chan W, Eberhardt R, Gardner M, Laiho K, Legge D, Magrane M, Pichler K, Poggioli D, Sehra H, Auchincloss A, Axelsen K, Blatter MC, Boutet E, Braconi-Quintaje S, Breuza L, Bridge A, Coudert E, Estreicher A, Famiglietti L, Ferro-Rojas S, Feuermann M, Gos A, Gruaz-Gumowski N, Hinz U, Hulo C, James J, Jimenez S, Jungo F, Keller G, Lemercier P, Lieberherr D, Masson P, Moinat M, Pedruzzi I, Poux S, Rivoire C, Roechert B, Schneider M, Stutz A, Sundaram S, Tognolli M, Bougueleret L, Argoud-Puy G, Cusin I, Duek-Roggli P, Xenarios I, Apweiler R

Abstract

The GO annotation dataset provided by the UniProt Consortium (GOA: http://www.ebi.ac.uk/GOA) is a comprehensive set of evidenced-based associations between terms from the Gene Ontology resource and UniProtKB proteins. Currently supplying over 100 million annotations to 11 million proteins in more than 360,000 taxa, this resource has increased 2-fold over the last 2 years and has benefited from a wealth of checks to improve annotation correctness and consistency as well as now supplying a greater information content enabled by GO Consortium annotation format developments. Detailed, manual GO annotations obtained from the curation of peer-reviewed papers are directly contributed by all UniProt curators and supplemented with manual and electronic annotations from 36 model organism and domain-focused scientific resources. The inclusion of high-quality, automatic annotation predictions ensures the UniProt GO annotation dataset supplies functional information to a wide range of proteins, including those from poorly characterized, non-model organism species. UniProt GO annotations are freely available in a range of formats accessible by both file downloads and web-based views. In addition, the introduction of a new, normalized file format in 2010 has made for easier handling of the complete UniProt-GOA data set.

MeSH Terms
Databases, Protein Molecular Sequence Annotation/standards Vocabulary, Controlled
Authors & Affiliations
55 authors, click to expand affiliations / ORCID
Dimmer Emily C
European Bioinformatics Institute, Wellcome Trust Genome Campus, Hinxton, Cambridge CB10 1SD, UK. edimmer@ebi.ac.uk
Huntley Rachael P
Alam-Faruque Yasmin
Sawford Tony
O'Donovan Claire
Martin Maria J
Bely Benoit
Browne Paul
Mun Chan Wei
Eberhardt Ruth
Gardner Michael
Laiho Kati
Legge Duncan
Magrane Michele
Pichler Klemens
Poggioli Diego
Sehra Harminder
Auchincloss Andrea
Axelsen Kristian
Blatter Marie-Claude
Boutet Emmanuel
Braconi-Quintaje Silvia
Breuza Lionel
Bridge Alan
Coudert Elizabeth
Estreicher Anne
Famiglietti Livia
Ferro-Rojas Serenella
Feuermann Marc
Gos Arnaud
Gruaz-Gumowski Nadine
Hinz Ursula
Hulo Chantal
James Janet
Jimenez Silvia
Jungo Florence
Keller Guillaume
Lemercier Phillippe
Lieberherr Damien
Masson Patrick
Moinat Madelaine
Pedruzzi Ivo
Poux Sylvain
Rivoire Catherine
Roechert Bernd
Schneider Michael
Stutz Andre
Sundaram Shyamala
Tognolli Michael
Bougueleret Lydie
Argoud-Puy Ghislaine
Cusin Isabelle
Duek-Roggli Paula
Xenarios Ioannis
Apweiler Rolf
References (25)
25 references, click to expand
  1. Gramene database: a hub for comparative plant genomics.
    Methods Mol Biol. 2011;678:247-75 PMID: 20931385
  2. The Renal Gene Ontology Annotation Initiative.
    Organogenesis. 2010 Apr-Jun;6(2):71-5 PMID: 20885853
  3. InterPro: the integrative protein signature database.
    Nucleic Acids Res. 2009 Jan;37(Database issue):D211-5 PMID: 18940856
  4. EnsemblCompara GeneTrees: Complete, duplication-aware phylogenetic trees in vertebrates.
    Genome Res. 2009 Feb;19(2):327-35 PMID: 19029536
  5. Improvements to cardiovascular gene ontology.
    Atherosclerosis. 2009 Jul;205(1):9-14 PMID: 19046747
  6. WormBase: a comprehensive resource for nematode research.
    Nucleic Acids Res. 2010 Jan;38(Database issue):D463-7 PMID: 19910365
  7. The IntAct molecular interaction database in 2010.
    Nucleic Acids Res. 2010 Jan;38(Database issue):D525-31 PMID: 19850723
  8. AgBase: supporting functional modeling in agricultural organisms.
    Nucleic Acids Res. 2011 Jan;39(Database issue):D497-506 PMID: 21075795
  9. The GOA database in 2009--an integrated Gene Ontology Annotation resource.
    Nucleic Acids Res. 2009 Jan;37(Database issue):D396-403 PMID: 18957448
  10. QuickGO: a web-based tool for Gene Ontology searching.
    Bioinformatics. 2009 Nov 15;25(22):3045-6 PMID: 19744993
  11. dictyBase update 2011: web 2.0 functionality and the initial steps towards a genome portal for the Amoebozoa.
    Nucleic Acids Res. 2011 Jan;39(Database issue):D620-4 PMID: 21087999
  12. The Integr8 project--a resource for genomic and proteomic data.
    In Silico Biol. 2005;5(2):179-85 PMID: 15972013
  13. Ongoing and future developments at the Universal Protein Resource.
    Nucleic Acids Res. 2011 Jan;39(Database issue):D214-9 PMID: 21051339
  14. What we can learn about Escherichia coli through application of Gene Ontology.
    Trends Microbiol. 2009 Jul;17(7):269-78 PMID: 19576778
  15. The International Protein Index: an integrated database for proteomics experiments.
    Proteomics. 2004 Jul;4(7):1985-8 PMID: 15221759
  16. The Gene Ontology: enhancements for 2011.
    Nucleic Acids Res. 2012 Jan;40(Database issue):D559-64 PMID: 22102568
  17. Chemical Entities of Biological Interest: an update.
    Nucleic Acids Res. 2010 Jan;38(Database issue):D249-54 PMID: 19854951
  18. The comprehensive microbial resource.
    Nucleic Acids Res. 2010 Jan;38(Database issue):D340-5 PMID: 19892825
  19. Reorganizing the protein space at the Universal Protein Resource (UniProt).
    Nucleic Acids Res. 2012 Jan;40(Database issue):D71-5 PMID: 22102590
  20. EcoCyc: a comprehensive database of Escherichia coli biology.
    Nucleic Acids Res. 2011 Jan;39(Database issue):D583-90 PMID: 21097882
  21. New tools at the Candida Genome Database: biochemical pathways and full-text literature search.
    Nucleic Acids Res. 2010 Jan;38(Database issue):D428-32 PMID: 19808938
  22. Logical development of the cell ontology.
    BMC Bioinformatics. 2011 Jan 05;12:6 PMID: 21208450
  23. The Gene Ontology in 2010: extensions and refinements.
    Nucleic Acids Res. 2010 Jan;38(Database issue):D331-5 PMID: 19920128
  24. The Plant-Associated Microbe Gene Ontology (PAMGO) Consortium: community development of new Gene Ontology terms describing biological processes involved in microbe-host interactions.
    BMC Microbiol. 2009 Feb 19;9 Suppl 1:S1 PMID: 19278549
  25. Formalization of taxon-based constraints to detect inconsistencies in annotation and ontology development.
    BMC Bioinformatics. 2010 Oct 25;11:530 PMID: 20973947
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
1362-4962
Published
2012-01-00
Epub
2011-00-28
Pages
D565-70
Language
English
Region
England
NLM ID
0411011
PMCID
PMC3245010
Subset
IM
Grants
NHGRI NIH HHS · 1U41HG006104-02 · United States
NHGRI NIH HHS · 3P41HG002273-09 · United States
British Heart Foundation · SP:07/007/23671 · United Kingdom
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com