Abstract
The iProClass database is an integrated resource that provides comprehensive family relationships and structural and functional features of proteins, with rich links to various databases. It is extended from ProClass, a protein family database that integrates PIR superfamilies and PROSITE motifs. The iProClass currently consists of more than 200,000 non-redundant PIR and SWISS-PROT proteins organized with more than 28,000 superfamilies, 2600 domains, 1300 motifs, 280 post-translational modification sites and links to more than 30 databases of protein families, structures, functions, genes, genomes, literature and taxonomy. Protein and family summary reports provide rich annotations, including membership information with length, taxonomy and keyword statistics, full family relationships, comprehensive enzyme and PDB cross-references and graphical feature display. The database facilitates classification-driven annotation for protein sequence databases and complete genomes, and supports structural and functional genomic research. The iProClass is implemented in Oracle 8i object-relational system and available for sequence search and report retrieval at http://pir.georgetown.edu/iproclass/.
MeSH Terms
Databases, Factual
Information Services
Internet
Proteins/classification
Authors & Affiliations
5 authors, click to expand affiliations / ORCID
Wu C H
Protein Information Resource, National Biomedical Research Foundation, Georgetown University Medical Center, 3900 Reservoir Road, NW Washington, DC 20007-2195, USA. wuc@nbrf.georgetown.edu
Xiao C
Hou Z
Huang H
Barker W C
References (16)
16 references, click to expand
-
The COG database: a tool for genome-scale analysis of protein functions and evolution.
Nucleic Acids Res. 2000 Jan 1;28(1):33-6
PMID: 10592175
-
PIR-ALN: a database of protein sequence alignments.
Bioinformatics. 1999 May;15(5):382-90
PMID: 10366659
-
The RESID database of protein structure modifications: 2000 update.
Nucleic Acids Res. 2000 Jan 1;28(1):209-11
PMID: 10592227
-
PRINTS-S: the database formerly known as PRINTS.
Nucleic Acids Res. 2000 Jan 1;28(1):225-7
PMID: 10592232
-
Increased coverage of protein families with the blocks database servers.
Nucleic Acids Res. 2000 Jan 1;28(1):228-30
PMID: 10592233
-
The Pfam protein families database.
Nucleic Acids Res. 2000 Jan 1;28(1):263-6
PMID: 10592242
-
ProDom and ProDom-CG: tools for protein domain analysis and whole genome comparisons.
Nucleic Acids Res. 2000 Jan 1;28(1):267-9
PMID: 10592243
-
ProClass protein family database.
Nucleic Acids Res. 2000 Jan 1;28(1):273-6
PMID: 10592245
-
Protein Information Resource: a community resource for expert annotation of protein data.
Nucleic Acids Res. 2001 Jan 1;29(1):29-32
PMID: 11125041
-
The MetaFam Server: a comprehensive protein family resource.
Nucleic Acids Res. 2001 Jan 1;29(1):49-51
PMID: 11125046
-
Superfamily classification in PIR-International Protein Sequence Database.
Methods Enzymol. 1996;266:59-71
PMID: 8743677
-
A protein class database organized with ProSite protein groups and PIR superfamilies.
J Comput Biol. 1996 Winter;3(4):547-61
PMID: 9018603
-
Gapped BLAST and PSI-BLAST: a new generation of protein database search programs.
Nucleic Acids Res. 1997 Sep 1;25(17):3389-402
PMID: 9254694
-
The PROSITE database, its status in 1999.
Nucleic Acids Res. 1999 Jan 1;27(1):215-9
PMID: 9847184
-
SCOP: a Structural Classification of Proteins database.
Nucleic Acids Res. 1999 Jan 1;27(1):254-6
PMID: 9847194
-
The SWISS-PROT protein sequence database and its supplement TrEMBL in 2000.
Nucleic Acids Res. 2000 Jan 1;28(1):45-8
PMID: 10592178