Abstract
The Protein Information Resource, in collaboration with the Munich Information Center for Protein Sequences (MIPS) and the Japan International Protein Information Database (JIPID), produces the most comprehensive and expertly annotated protein sequence database in the public domain, the PIR-International Protein Sequence Database. To provide timely and high quality annotation and promote database interoperability, the PIR-International employs rule-based and classification-driven procedures based on controlled vocabulary and standard nomenclature and includes status tags to distinguish experimentally determined from predicted protein features. The database contains about 200,000 non-redundant protein sequences, which are classified into families and superfamilies and their domains and motifs identified. Entries are extensively cross-referenced to other sequence, classification, genome, structure and activity databases. The PIR web site features search engines that use sequence similarity and database annotation to facilitate the analysis and functional identification of proteins. The PIR-Inter-national databases and search tools are accessible on the PIR web site at http://pir.georgetown.edu/ and at the MIPS web site at http://www.mips.biochem.mpg.de. The PIR-International Protein Sequence Database and other files are also available by FTP.
MeSH Terms
Computational Biology
Databases, Factual
Information Services
Internet
Proteins/classification,genetics
Terminology as Topic
Authors & Affiliations
14 authors, click to expand affiliations / ORCID
Barker W C
National Biomedical Research Foundation, 3900 Reservoir Road, LR-3, NW, Washington, DC 20007, USA. pirmail@nbrf.georgetown.edu
Garavelli J S
Hou Z
Huang H
Ledley R S
McGarvey P B
Mewes H W
Orcutt B C
Pfeiffer F
Tsugita A
Vinayaka C R
Xiao C
Yeh L S
Wu C
References (13)
13 references, click to expand
-
The COG database: a tool for genome-scale analysis of protein functions and evolution.
Nucleic Acids Res. 2000 Jan 1;28(1):33-6
PMID: 10592175
-
The Pfam protein families database.
Nucleic Acids Res. 2000 Jan 1;28(1):263-6
PMID: 10592242
-
ProClass protein family database.
Nucleic Acids Res. 2000 Jan 1;28(1):273-6
PMID: 10592245
-
PIR: a new resource for bioinformatics.
Bioinformatics. 2000 Mar;16(3):290-1
PMID: 10869023
-
iProClass: an integrated, comprehensive and annotated protein classification database.
Nucleic Acids Res. 2001 Jan 1;29(1):52-4
PMID: 11125047
-
The RESID Database of protein structure modifications and the NRL-3D Sequence-Structure Database.
Nucleic Acids Res. 2001 Jan 1;29(1):199-201
PMID: 11125090
-
PIR-ALN: a database of protein sequence alignments.
Bioinformatics. 1999 May;15(5):382-90
PMID: 10366659
-
CLUSTAL W: improving the sensitivity of progressive multiple sequence alignment through sequence weighting, position-specific gap penalties and weight matrix choice.
Nucleic Acids Res. 1994 Nov 11;22(22):4673-80
PMID: 7984417
-
Maximum discrimination hidden Markov models of sequence consensus.
J Comput Biol. 1995 Spring;2(1):9-23
PMID: 7497123
-
Superfamily classification in PIR-International Protein Sequence Database.
Methods Enzymol. 1996;266:59-71
PMID: 8743677
-
Gapped BLAST and PSI-BLAST: a new generation of protein database search programs.
Nucleic Acids Res. 1997 Sep 1;25(17):3389-402
PMID: 9254694
-
The PROSITE database, its status in 1999.
Nucleic Acids Res. 1999 Jan 1;27(1):215-9
PMID: 9847184
-
Improved tools for biological sequence comparison.
Proc Natl Acad Sci U S A. 1988 Apr;85(8):2444-8
PMID: 3162770