Home LiteratureArticle Details
PMID: 11590095 Published · ppublish English Journal Article Research Support, U.S. Gov't, P.H.S.

A non-parametric approach to translating gene region heterogeneity associated with phenotype into location heterogeneity.

Bioinformatics (Oxford, England) ·Vol. 17 ·No. 9 ·2001-09-00 ·Pages 775-90

Kowalski J

Abstract

The analysis of genetic data poses statistical problems in the form of high dimensionality with small sample sizes. The construction of a composite gene region (sequence pair) heterogeneity measure is one technique for reducing the dimensionality of the problem. This approach however is not without cost, since the contribution of locations to observed gene region differences between groups becomes entangled in this summary measure. This is problematic since it is of scientific interest to identify locations that together depict phenotype. A method is proposed for relating observed gene region heterogeneity back to the location level. In the spirit of a factor analysis-type setting, the approach focuses on identifying a latent variable structure among locations to explain within and between group genetic differences associated with phenotype. The method is flexible for identifying either the additive contribution from individual locations or the additive contribution from a group of locations, to observed gene region heterogeneity, depending upon the weighting scheme used in constructing a gene region heterogeneity measure. The approach is illustrated with clinical trial data, where the problem of altered HIV drug susceptibility is examined through characterizing location contributions to HIV protease gene region differences associated with a phenotypic treatment response. The Splus (MathSoft, Inc. S-Plus 2000, Seattle, WA, 1999) developed menu-driven functions for obtaining results, GENE_ S (J.Kowalski, Harvard School of Public Health, Boston, MA 2001), is available from the author upon request.

MeSH Terms
Clinical Trials, Phase II as Topic/statistics & numerical data Factor Analysis, Statistical Genes, Viral/genetics Genetic Heterogeneity Genetic Markers/genetics Genome, Viral HIV Protease/chemistry,genetics HIV-1/enzymology,genetics Multicenter Studies as Topic/statistics & numerical data Phenotype Protein Biosynthesis Randomized Controlled Trials as Topic/statistics & numerical data Sequence Analysis, DNA/methods,statistics & numerical data Statistics, Nonparametric Viral Structural Proteins/genetics
Chemicals
Genetic Markers Viral Structural Proteins HIV Protease
Authors & Affiliations
1 authors, click to expand affiliations / ORCID
Kowalski J
Department of Biostatistics, Harvard School of Public Health, Boston, MA 02115, USA.
Article Info
Journal
Bioinformatics (Oxford, England)
Abbr.
Bioinformatics
ISSN
1367-4803
Published
2001-09-00
Pages
775-90
Language
English
Region
England
NLM ID
9808944
Subset
IM
Grants
NIAID NIH HHS · T32-AI07358 · United States
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com