Abstract
The DBAli tools use a comprehensive set of structural alignments in the DBAli database to leverage the structural information deposited in the Protein Data Bank (PDB). These tools include (i) the DBAlit program that allows users to input the 3D coordinates of a protein structure for comparison by MAMMOTH against all chains in the PDB; (ii) the AnnoLite and AnnoLyze programs that annotate a target structure based on its stored relationships to other structures; (iii) the ModClus program that clusters structures by sequence and structure similarities; (iv) the ModDom program that identifies domains as recurrent structural fragments and (v) an implementation of the COMPARER method in the SALIGN command in MODELLER that creates a multiple structure alignment for a set of related protein structures. Thus, the DBAli tools, which are freely accessible via the World Wide Web at http://salilab.org/DBAli/, allow users to mine the protein structure space by establishing relationships between protein structures and their functions.
MeSH Terms
Algorithms
Amino Acid Sequence
Computational Biology/methods
Data Interpretation, Statistical
Databases, Protein
Internet
Molecular Sequence Data
Protein Conformation
Proteins/chemistry,classification,metabolism
Pseudomonas aeruginosa/metabolism
Sequence Alignment/methods
Sequence Analysis, Protein/methods
Sequence Homology, Amino Acid
Software
Structure-Activity Relationship
Authors & Affiliations
9 authors, click to expand affiliations / ORCID
Marti-Renom Marc A
Structural Genomics Unit, and California Institute for Quantitative Biomedical Research, University of California at San Francisco, San Francisco, CA 94158-2330, USA. mmarti@cipf.es
Pieper Ursula
Madhusudhan M S
Rossi Andrea
Eswar Narayanan
Davis Fred P
Al-Shahrour Fátima
Dopazo Joaquín
Sali Andrej
References (25)
25 references, click to expand
-
The CATH Domain Structure Database and related resources Gene3D and DHS provide comprehensive domain family information for genome analysis.
Nucleic Acids Res. 2005 Jan 1;33(Database issue):D247-51
PMID: 15608188
-
InterPro, progress and status in 2005.
Nucleic Acids Res. 2005 Jan 1;33(Database issue):D201-5
PMID: 15608177
-
Inference of protein function from protein structure.
Structure. 2005 Jan;13(1):121-30
PMID: 15642267
-
PIBASE: a comprehensive database of structurally defined protein interfaces.
Bioinformatics. 2005 May 1;21(9):1901-7
PMID: 15657096
-
ProFunc: a server for predicting protein function from 3D structure.
Nucleic Acids Res. 2005 Jul 1;33(Web Server issue):W89-93
PMID: 15980588
-
MODBASE: a database of annotated comparative protein structure models and associated resources.
Nucleic Acids Res. 2006 Jan 1;34(Database issue):D291-5
PMID: 16381869
-
The AnnoLite and AnnoLyze programs for comparative annotation of protein structures.
BMC Bioinformatics. 2007;8 Suppl 4:S4
PMID: 17570147
-
The Protein Data Bank.
Nucleic Acids Res. 2000 Jan 1;28(1):235-42
PMID: 10592235
-
The ENZYME database in 2000.
Nucleic Acids Res. 2000 Jan 1;28(1):304-5
PMID: 10592255
-
Gene ontology: tool for the unification of biology. The Gene Ontology Consortium.
Nat Genet. 2000 May;25(1):25-9
PMID: 10802651
-
Completeness in structural genomics.
Nat Struct Biol. 2001 Jun;8(6):559-66
PMID: 11373627
-
DBAli: a database of protein structure alignments.
Bioinformatics. 2001 Aug;17(8):746-7
PMID: 11524379
-
LigBase: a database of families of aligned ligand binding sites in known protein sequences and structures.
Bioinformatics. 2002 Jan;18(1):200-1
PMID: 11836232
-
MAMMOTH (matching molecular models obtained from theory): an automated method for model comparison.
Protein Sci. 2002 Nov;11(11):2606-21
PMID: 12381844
-
The Gene Ontology Annotation (GOA) project: implementation of GO in SWISS-PROT, TrEMBL, and InterPro.
Genome Res. 2003 Apr;13(4):662-72
PMID: 12654719
-
From protein structure to biochemical function?
J Struct Funct Genomics. 2003;4(2-3):167-77
PMID: 14649301
-
The Pfam protein families database.
Nucleic Acids Res. 2004 Jan 1;32(Database issue):D138-41
PMID: 14681378
-
SCOP database in 2004: refinements integrate structure and sequence family data.
Nucleic Acids Res. 2004 Jan 1;32(Database issue):D226-9
PMID: 14681400
-
FatiGO: a web tool for finding significant associations of Gene Ontology terms with groups of genes.
Bioinformatics. 2004 Mar 1;20(4):578-80
PMID: 14990455
-
Automated prediction of protein function and detection of functional sites from structure.
Proc Natl Acad Sci U S A. 2004 Oct 12;101(41):14754-9
PMID: 15456910
-
Definition of general topological equivalence in protein structures. A procedure involving comparison of properties and relationships through simulated annealing and dynamic programming.
J Mol Biol. 1990 Mar 20;212(2):403-28
PMID: 2181150
-
Gapped BLAST and PSI-BLAST: a new generation of protein database search programs.
Nucleic Acids Res. 1997 Sep 1;25(17):3389-402
PMID: 9254694
-
GenTHREADER: an efficient and reliable protein fold recognition method for genomic sequences.
J Mol Biol. 1999 Apr 9;287(4):797-815
PMID: 10191147
-
Structural genomics: beyond the human genome project.
Nat Genet. 1999 Oct;23(2):151-7
PMID: 10508510
-
E-MSD: an integrated data resource for bioinformatics.
Nucleic Acids Res. 2005 Jan 1;33(Database issue):D262-5
PMID: 15608192