Home LiteratureArticle Details
PMID: 18948282 Published · ppublish English Journal Article Research Support, N.I.H., Extramural Research Support, Non-U.S. Gov't Research Support, U.S. Gov't, Non-P.H.S.

MODBASE, a database of annotated comparative protein structure models and associated resources.

Nucleic acids research ·Vol. 37 ·No. Database issue ·2009-01-00 ·Pages D347-54

Pieper U, Eswar N, Webb BM, Eramian D, Kelly L, Barkan DT, Carter H, Mankoo P, Karchin R, Marti-Renom MA, Davis FP, Sali A

Abstract

MODBASE (http://salilab.org/modbase) is a database of annotated comparative protein structure models. The models are calculated by MODPIPE, an automated modeling pipeline that relies primarily on MODELLER for fold assignment, sequence-structure alignment, model building and model assessment (http:/salilab.org/modeller). MODBASE currently contains 5,152,695 reliable models for domains in 1,593,209 unique protein sequences; only models based on statistically significant alignments and/or models assessed to have the correct fold are included. MODBASE also allows users to calculate comparative models on demand, through an interface to the MODWEB modeling server (http://salilab.org/modweb). Other resources integrated with MODBASE include databases of multiple protein structure alignments (DBAli), structurally defined ligand binding sites (LIGBASE), predicted ligand binding sites (AnnoLyze), structurally defined binary domain interfaces (PIBASE) and annotated single nucleotide polymorphisms and somatic mutations found in human proteins (LS-SNP, LS-Mut). MODBASE models are also available through the Protein Model Portal (http://www.proteinmodelportal.org/).

MeSH Terms
Databases, Protein Genomics Humans Ligands Models, Molecular Mutation Polymorphism, Single Nucleotide Protein Folding Protein Interaction Domains and Motifs Protein Structure, Tertiary Proteins/genetics Structural Homology, Protein User-Computer Interface
Chemicals
Ligands Proteins
Authors & Affiliations
12 authors, click to expand affiliations / ORCID
Pieper Ursula
Department of Bioengineering and Therapeutic Sciences, Department of Pharmaceutical Chemistry, University of California at San Francisco, 1700 4th Street, San Francisco, CA 94158, USA.
Eswar Narayanan
Webb Ben M
Eramian David
Kelly Libusha
Barkan David T
Carter Hannah
Mankoo Parminder
Karchin Rachel
Marti-Renom Marc A
Davis Fred P
Sali Andrej
References (54)
54 references, click to expand
  1. Natural variation in human membrane transporter genes reveals evolutionary and functional constraints.
    Proc Natl Acad Sci U S A. 2003 May 13;100(10):5896-901 PMID: 12719533
  2. Protein structure prediction and structural genomics.
    Science. 2001 Oct 5;294(5540):93-6 PMID: 11588250
  3. The AnnoLite and AnnoLyze programs for comparative annotation of protein structures.
    BMC Bioinformatics. 2007;8 Suppl 4:S4 PMID: 17570147
  4. Finding cures for tropical diseases: is open source an answer?
    PLoS Med. 2004 Dec;1(3):e56 PMID: 15630466
  5. A composite score for predicting errors in protein structure models.
    Protein Sci. 2006 Jul;15(7):1653-66 PMID: 16751606
  6. Structural genomics and its importance for gene function analysis.
    Nat Biotechnol. 2000 Mar;18(3):283-7 PMID: 10700142
  7. Tools for comparative protein structure modeling and analysis.
    Nucleic Acids Res. 2003 Jul 1;31(13):3375-80 PMID: 12824331
  8. An integrated genomic analysis of human glioblastoma multiforme.
    Science. 2008 Sep 26;321(5897):1807-12 PMID: 18772396
  9. Learning from the genome sequence of Mycobacterium tuberculosis H37Rv.
    FEBS Lett. 1999 Jun 4;452(1-2):7-10 PMID: 10376668
  10. The human ATP-binding cassette (ABC) transporter superfamily.
    Genome Res. 2001 Jul;11(7):1156-66 PMID: 11435397
  11. dbSNP: the NCBI database of genetic variation.
    Nucleic Acids Res. 2001 Jan 1;29(1):308-11 PMID: 11125122
  12. Host pathogen protein interactions predicted by comparative modeling.
    Protein Sci. 2007 Dec;16(12):2585-96 PMID: 17965183
  13. Statistical potentials for fold assessment.
    Protein Sci. 2002 Feb;11(2):430-48 PMID: 11790853
  14. MAMMOTH (matching molecular models obtained from theory): an automated method for model comparison.
    Protein Sci. 2002 Nov;11(11):2606-21 PMID: 12381844
  15. Molecular recognition. Conformational analysis of limited proteolytic sites and serine proteinase protein inhibitors.
    J Mol Biol. 1991 Jul 20;220(2):507-30 PMID: 1856871
  16. Dictionary of protein secondary structure: pattern recognition of hydrogen-bonded and geometrical features.
    Biopolymers. 1983 Dec;22(12):2577-637 PMID: 6667333
  17. The Universal Protein Resource (UniProt).
    Nucleic Acids Res. 2005 Jan 1;33(Database issue):D154-9 PMID: 15608167
  18. CryptoDB: a Cryptosporidium bioinformatics resource update.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D419-22 PMID: 16381902
  19. OrthoMCL-DB: querying a comprehensive multi-species collection of ortholog groups.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D363-8 PMID: 16381887
  20. LS-SNP: large-scale annotation of coding non-synonymous SNPs based on multiple information sources.
    Bioinformatics. 2005 Jun 15;21(12):2814-20 PMID: 15827081
  21. UCSF Chimera--a visualization system for exploratory research and analysis.
    J Comput Chem. 2004 Oct;25(13):1605-12 PMID: 15264254
  22. GenBank.
    Nucleic Acids Res. 2008 Jan;36(Database issue):D25-30 PMID: 18073190
  23. Online Mendelian Inheritance in Man (OMIM), a knowledgebase of human genes and genetic disorders.
    Nucleic Acids Res. 2005 Jan 1;33(Database issue):D514-7 PMID: 15608251
  24. Structure-based activity prediction for an enzyme of unknown function.
    Nature. 2007 Aug 16;448(7155):775-9 PMID: 17603473
  25. Draft genome of the filarial nematode parasite Brugia malayi.
    Science. 2007 Sep 21;317(5845):1756-60 PMID: 17885136
  26. Protein complex compositions predicted by structural similarity.
    Nucleic Acids Res. 2006;34(10):2943-52 PMID: 16738133
  27. Statistical potential for assessment and prediction of protein structures.
    Protein Sci. 2006 Nov;15(11):2507-24 PMID: 17075131
  28. PIBASE: a comprehensive database of structurally defined protein interfaces.
    Bioinformatics. 2005 May 1;21(9):1901-7 PMID: 15657096
  29. Core signaling pathways in human pancreatic cancers revealed by global genomic analyses.
    Science. 2008 Sep 26;321(5897):1801-6 PMID: 18772397
  30. ToxoDB: an integrated Toxoplasma gondii database resource.
    Nucleic Acids Res. 2008 Jan;36(Database issue):D553-6 PMID: 18003657
  31. The role of protein structure in genomics.
    FEBS Lett. 2000 Jun 30;476(1-2):98-102 PMID: 10878259
  32. DBAli tools: mining the protein structure space.
    Nucleic Acids Res. 2007 Jul;35(Web Server issue):W393-7 PMID: 17478513
  33. Comparative protein structure modeling using Modeller.
    Curr Protoc Bioinformatics. 2006 Oct;Chapter 5:Unit 5.6 PMID: 18428767
  34. The UCSC Known Genes.
    Bioinformatics. 2006 May 1;22(9):1036-46 PMID: 16500937
  35. Ensembl 2008.
    Nucleic Acids Res. 2008 Jan;36(Database issue):D707-14 PMID: 18000006
  36. Raster3D: photorealistic molecular graphics.
    Methods Enzymol. 1997;277:505-24 PMID: 18488322
  37. Alignment of protein sequences by their profiles.
    Protein Sci. 2004 Apr;13(4):1071-87 PMID: 15044736
  38. Global sequencing of proteolytic cleavage sites in apoptosis by specific labeling of protein N termini.
    Cell. 2008 Sep 5;134(5):866-76 PMID: 18722006
  39. The RCSB Protein Data Bank: a redesigned query system and relational database based on the mmCIF schema.
    Nucleic Acids Res. 2005 Jan 1;33(Database issue):D233-7 PMID: 15608185
  40. Utility of homology models in the drug discovery process.
    Drug Discov Today. 2004 Aug 1;9(15):659-69 PMID: 15279849
  41. High-throughput computational and experimental techniques in structural genomics.
    Genome Res. 2004 Oct;14(10B):2145-54 PMID: 15489337
  42. All are not equal: a benchmark of different homology modeling programs.
    Protein Sci. 2005 May;14(5):1315-27 PMID: 15840834
  43. The Universal Protein Resource (UniProt): an expanding universe of protein information.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D187-91 PMID: 16381842
  44. Database resources of the National Center for Biotechnology Information.
    Nucleic Acids Res. 2008 Jan;36(Database issue):D13-21 PMID: 18045790
  45. MODBASE: a database of annotated comparative protein structure models and associated resources.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D291-5 PMID: 16381869
  46. Expectations from structural genomics.
    Protein Sci. 2000 Jan;9(1):197-200 PMID: 10739263
  47. GeneDB: a resource for prokaryotic and eukaryotic organisms.
    Nucleic Acids Res. 2004 Jan 1;32(Database issue):D339-43 PMID: 14681429
  48. Leveraging enzyme structure-function relationships for functional inference and experimental design: the structure-function linkage database.
    Biochemistry. 2006 Feb 28;45(8):2545-55 PMID: 16489747
  49. Identification of common molecular subsequences.
    J Mol Biol. 1981 Mar 25;147(1):195-7 PMID: 7265238
  50. Gapped BLAST and PSI-BLAST: a new generation of protein database search programs.
    Nucleic Acids Res. 1997 Sep 1;25(17):3389-402 PMID: 9254694
  51. DBAli: a database of protein structure alignments.
    Bioinformatics. 2001 Aug;17(8):746-7 PMID: 11524379
  52. Comparative protein modelling by satisfaction of spatial restraints.
    J Mol Biol. 1993 Dec 5;234(3):779-815 PMID: 8254673
  53. Comparative protein structure modeling using MODELLER.
    Curr Protoc Protein Sci. 2007 Nov;Chapter 2:Unit 2.9 PMID: 18429317
  54. LigBase: a database of families of aligned ligand binding sites in known protein sequences and structures.
    Bioinformatics. 2002 Jan;18(1):200-1 PMID: 11836232
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
1362-4962
Published
2009-01-00
Epub
2008-00-23
Pages
D347-54
Language
English
Region
England
NLM ID
0411011
PMCID
PMC2686492
Subset
IM
Grants
NIGMS NIH HHS · U01 GM61390 · United States
NIGMS NIH HHS · R01 GM54762 · United States
NIGMS NIH HHS · GM08284 · United States
NIGMS NIH HHS · P01 GM71790 · United States
NIGMS NIH HHS · U54 GM074945 · United States
NIGMS NIH HHS · U54 GM074929 · United States
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com