Home LiteratureArticle Details
PMID: 9199410 Published · ppublish English Journal Article

Recognition of analogous and homologous protein folds: analysis of sequence and structure conservation.

Journal of molecular biology ·Vol. 269 ·No. 3 ·1997-06-13 ·Pages 423-39

Russell RB, Saqi MA, Sayle RA, Bates PA, Sternberg MJ

Abstract

An analysis was performed on 335 pairs of structurally aligned proteins derived from the structural classification of proteins (SCOP http://scop.mrc-lmb.cam.ac.uk/scop/) database. These similarities were divided into analogues, defined as proteins with similar three-dimensional structures (same SCOP fold classification) but generally with different functions and little evidence of a common ancestor (different SCOP superfamily classification). Homologues were defined as pairs of similar structures likely to be the result of evolutionary divergence (same superfamily) and were divided into remote, medium and close sub-divisions based on the percentage sequence identity. Particular attention was paid to the differences between analogues and remote homologues, since both types of similarities are generally undetectable by sequence comparison and their detection is the aim of fold recognition methods. Distributions of sequence identities and substitution matrices suggest a higher degree of sequence similarity in remote homologues than in analogues. Matrices for remote homologues show similarity to existing mutation matrices, providing some validity for their use in previously described fold recognition methods. In contrast, matrices derived from analogous proteins show little conservation of amino acid properties beyond broad conservation of hydrophobic or polar character. Secondary structure and accessibility were more conserved on average in remote homologues than in analogues, though there was no apparent difference in the root-mean-square deviation between these two types of similarities. Alignments of remote homologues and analogues show a similar number of gaps, openings (one or more sequential gaps) and inserted/deleted secondary structure elements, and both generally contain more gaps/openings/deleted secondary structure elements than medium and close homologues. These results suggest that gap parameters for fold recognition should be more lenient than those used in sequence comparison. Parameters were derived from the analogue and remote homologue datasets for potential used in fold recognition methods. Implications for protein fold recognition and evolution are discussed.

MeSH Terms
Computer Simulation DNA Transposable Elements Databases, Factual Models, Molecular Mutation Protein Folding Proteins/chemistry,genetics Sequence Alignment Sequence Analysis/methods Sequence Deletion Sequence Homology, Amino Acid
Chemicals
DNA Transposable Elements Proteins
Authors & Affiliations
5 authors, click to expand affiliations / ORCID
Russell R B
Biomolecular Modelling Laboratory, Imperial Cancer Research Fund, Lincoln's Inn Fields, London, UK.
Saqi M A
Sayle R A
Bates P A
Sternberg M J
Article Info
Journal
Journal of molecular biology
Abbr.
J Mol Biol
ISSN
0022-2836
Published
1997-06-13
Pages
423-39
Language
English
Region
England
NLM ID
2985088R
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com