Home LiteratureArticle Details
PMID: 10592242 Published · ppublish English Journal Article

The Pfam protein families database.

Nucleic acids research ·Vol. 28 ·No. 1 ·2000-01-01 ·Pages 263-6

Bateman A, Birney E, Durbin R, Eddy SR, Howe KL, Sonnhammer EL

Abstract

Pfam is a large collection of protein multiple sequence alignments and profile hidden Markov models. Pfam is available on the WWW in the UK at http://www.sanger.ac.uk/Software/Pfam/, in Sweden at http://www.cgr.ki.se/Pfam/ and in the US at http://pfam.wustl.edu/. The latest version (4.3) of Pfam contains 1815 families. These Pfam families match 63% of proteins in SWISS-PROT 37 and TrEMBL 9. For complete genomes Pfam currently matches up to half of the proteins. Genomic DNA can be directly searched against the Pfam library using the Wise2 package.

MeSH Terms
Databases, Factual Genome Information Storage and Retrieval Internet Proteins/chemistry Quality Control
Chemicals
Proteins
Authors & Affiliations
6 authors, click to expand affiliations / ORCID
Bateman A
The Sanger Centre, Wellcome Trust Genome Campus, Hinxton, Cambridge CB10 1SA, UK. agb@sanger.ac.uk
Birney E
Durbin R
Eddy S R
Howe K L
Sonnhammer E L
References (12)
12 references, click to expand
  1. Sources of systematic error in functional annotation of genomes: domain rearrangement, non-orthologous gene displacement and operon disruption.
    In Silico Biol. 1998;1(1):55-67 PMID: 11471243
  2. The Protein Data Bank: a computer-based archival file for macromolecular structures.
    J Mol Biol. 1977 May 25;112(3):535-42 PMID: 875032
  3. Dictionary of protein secondary structure: pattern recognition of hydrogen-bonded and geometrical features.
    Biopolymers. 1983 Dec;22(12):2577-637 PMID: 6667333
  4. Basic local alignment search tool.
    J Mol Biol. 1990 Oct 5;215(3):403-10 PMID: 2231712
  5. Modular arrangement of proteins as inferred from analysis of homology.
    Protein Sci. 1994 Mar;3(3):482-92 PMID: 8019419
  6. Genome sequence of the nematode C. elegans: a platform for investigating biology.
    Science. 1998 Dec 11;282(5396):2012-8 PMID: 9851916
  7. RASMOL: biomolecular graphics for all.
    Trends Biochem Sci. 1995 Sep;20(9):374 PMID: 7482707
  8. Pfam: a comprehensive database of protein domain families based on seed alignments.
    Proteins. 1997 Jul;28(3):405-20 PMID: 9223186
  9. Dynamite: a flexible code generating language for dynamic programming methods used in sequence comparison.
    Proc Int Conf Intell Syst Mol Biol. 1997;5:56-64 PMID: 9322016
  10. Pfam 3.1: 1313 multiple alignments and profile HMMs match the majority of proteins.
    Nucleic Acids Res. 1999 Jan 1;27(1):260-2 PMID: 9847196
  11. Recent improvements of the ProDom database of protein domain families.
    Nucleic Acids Res. 1999 Jan 1;27(1):263-7 PMID: 9847197
  12. CLUSTAL W: improving the sensitivity of progressive multiple sequence alignment through sequence weighting, position-specific gap penalties and weight matrix choice.
    Nucleic Acids Res. 1994 Nov 11;22(22):4673-80 PMID: 7984417
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
0305-1048
Published
2000-01-01
Pages
263-6
Language
English
Region
England
NLM ID
0411011
PMCID
PMC102420
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com