A Simple Loglinear Model for Haplotype Effects in a Case-Control Study Involving Two Unphased Genotypes
-
Stuart G. Baker
Because haplotypes may parsimoniously summarize the effect of genes on disease, there is great interest in using haplotypes in case-control studies of unphased genotype data. Previous methods for investigating haplotypes effects in case-control studies have not allowed for both of the following two scenarios that could have a large impact on results (i) departures from Hardy-Weinberg equilibrium in controls as well as cases, and (ii) an interactive effect of haplotypes and environmental covariates on the probability of disease. A new method is proposed that generalizes the model of Epstein and Satten to incorporate both (i) and (ii). Computations are relatively simple involving a single loglinear design matrix for parameters modeling the distribution of haplotype frequencies in controls, parameters modeling the effect of haplotypes and covariate-haplotype interactions on disease, and nuisance parameters required for correct inference. Based on simulations with realistic sample sizes, the method is recommended with data from two genotypes, a recessive or dominant model linking haplotypes to disease, and estimates of haplotype effects among haplotypes with a frequency greater than 10%. The methodology is most useful with candidate genotype pairs or for searching through pairs of genotypes when scenarios (i) and (ii) are likely. An example without a covariate illustrates the importance of modeling a departure from Hardy-Weinberg equilibrium in controls.
©2011 Walter de Gruyter GmbH & Co. KG, Berlin/Boston
Articles in the same Issue
- Article
- Estimating Motifs Under Order Restrictions
- Reproducible Research: A Bioinformatics Case Study
- Generalized Rank Tests for Replicated Microarray Data
- Stepwise Normalization of Two-Channel Spotted Microarrays
- Comparing Automatic and Manual Image Processing in FLARE Assay Analysis for Colon Carcinogenesis
- Pixel-level Signal Modelling with Spatial Correlation for Two-Colour Microarrays
- Empirical Bayes Microarray ANOVA and Grouping Cell Lines by Equal Expression Levels
- Multiple Testing and Data Adaptive Regression: An Application to HIV-1 Sequence Data.
- Early Diagnostic Marker Panel Determination for Microarray Based Clinical Studies
- Prediction of Missing Values in Microarray and Use of Mixed Models to Evaluate the Predictors
- Combined Association and Linkage Analysis for General Pedigrees and Genetic Models
- Incorporating Biological Information as a Prior in an Empirical Bayes Approach to Analyzing Microarray Data
- The Relative Inefficiency of Sequence Weights Approaches in Determining a Nucleotide Position Weight Matrix
- A Simple Loglinear Model for Haplotype Effects in a Case-Control Study Involving Two Unphased Genotypes
- Extension of the SIMLA Package for Generating Pedigrees with Complex Inheritance Patterns: Environmental Covariates, Gene-Gene and Gene-Environment Interaction
- Error Distribution for Gene Expression Data
- A General Framework for Weighted Gene Co-Expression Network Analysis
- Statistical Inference in Evolutionary Models of DNA Sequences via the EM Algorithm
- Comparing Bacterial DNA Microarray Fingerprints
- Continuous Covariates in Genetic Association Studies of Case-Parent Triads: Gene and Gene-Environment Interaction Effects, Population Stratification, and Power Analysis
- Robust Remote Homology Detection by Feature Based Profile Hidden Markov Models
- Empirical Bayes Estimation of a Sparse Vector of Gene Expression Changes
- Hierarchical Inverse Gaussian Models and Multiple Testing: Application to Gene Expression Data
- FADO: A Statistical Method to Detect Favored or Avoided Distances between Occurrences of Motifs using the Hawkes' Model
- Prediction of Genomewide Conserved Epitope Profiles of HIV-1: Classifier Choice and Peptide Representation
- Fold-Change Estimation of Differentially Expressed Genes using Mixture Mixed-Model
- Test on the Structure of Biological Sequences via Chaos Game Representation
- Reverse Engineering Galactose Regulation in Yeast through Model Selection
- Empirical Bayes and Resampling Based Multiple Testing Procedure Controlling Tail Probability of the Proportion of False Positives.
- Weighted Analysis of Paired Microarray Experiments
- A Probabilistic Approach to Large-Scale Association Scans: A Semi-Bayesian Method to Detect Disease-Predisposing Alleles
- A Shrinkage Approach to Large-Scale Covariance Matrix Estimation and Implications for Functional Genomics
- Structured Antedependence Models for Functional Mapping of Multiple Longitudinal Traits
- Correlation Between Gene Expression Levels and Limitations of the Empirical Bayes Methodology for Finding Differentially Expressed Genes
- Bayesian Statistical Studies of the Ramachandran Distribution
- On Reference Designs For Microarray Experiments
- Computing Asymptotic Power and Sample Size for Case-Control Genetic Association Studies in the Presence of Phenotype and/or Genotype Misclassification Errors
Articles in the same Issue
- Article
- Estimating Motifs Under Order Restrictions
- Reproducible Research: A Bioinformatics Case Study
- Generalized Rank Tests for Replicated Microarray Data
- Stepwise Normalization of Two-Channel Spotted Microarrays
- Comparing Automatic and Manual Image Processing in FLARE Assay Analysis for Colon Carcinogenesis
- Pixel-level Signal Modelling with Spatial Correlation for Two-Colour Microarrays
- Empirical Bayes Microarray ANOVA and Grouping Cell Lines by Equal Expression Levels
- Multiple Testing and Data Adaptive Regression: An Application to HIV-1 Sequence Data.
- Early Diagnostic Marker Panel Determination for Microarray Based Clinical Studies
- Prediction of Missing Values in Microarray and Use of Mixed Models to Evaluate the Predictors
- Combined Association and Linkage Analysis for General Pedigrees and Genetic Models
- Incorporating Biological Information as a Prior in an Empirical Bayes Approach to Analyzing Microarray Data
- The Relative Inefficiency of Sequence Weights Approaches in Determining a Nucleotide Position Weight Matrix
- A Simple Loglinear Model for Haplotype Effects in a Case-Control Study Involving Two Unphased Genotypes
- Extension of the SIMLA Package for Generating Pedigrees with Complex Inheritance Patterns: Environmental Covariates, Gene-Gene and Gene-Environment Interaction
- Error Distribution for Gene Expression Data
- A General Framework for Weighted Gene Co-Expression Network Analysis
- Statistical Inference in Evolutionary Models of DNA Sequences via the EM Algorithm
- Comparing Bacterial DNA Microarray Fingerprints
- Continuous Covariates in Genetic Association Studies of Case-Parent Triads: Gene and Gene-Environment Interaction Effects, Population Stratification, and Power Analysis
- Robust Remote Homology Detection by Feature Based Profile Hidden Markov Models
- Empirical Bayes Estimation of a Sparse Vector of Gene Expression Changes
- Hierarchical Inverse Gaussian Models and Multiple Testing: Application to Gene Expression Data
- FADO: A Statistical Method to Detect Favored or Avoided Distances between Occurrences of Motifs using the Hawkes' Model
- Prediction of Genomewide Conserved Epitope Profiles of HIV-1: Classifier Choice and Peptide Representation
- Fold-Change Estimation of Differentially Expressed Genes using Mixture Mixed-Model
- Test on the Structure of Biological Sequences via Chaos Game Representation
- Reverse Engineering Galactose Regulation in Yeast through Model Selection
- Empirical Bayes and Resampling Based Multiple Testing Procedure Controlling Tail Probability of the Proportion of False Positives.
- Weighted Analysis of Paired Microarray Experiments
- A Probabilistic Approach to Large-Scale Association Scans: A Semi-Bayesian Method to Detect Disease-Predisposing Alleles
- A Shrinkage Approach to Large-Scale Covariance Matrix Estimation and Implications for Functional Genomics
- Structured Antedependence Models for Functional Mapping of Multiple Longitudinal Traits
- Correlation Between Gene Expression Levels and Limitations of the Empirical Bayes Methodology for Finding Differentially Expressed Genes
- Bayesian Statistical Studies of the Ramachandran Distribution
- On Reference Designs For Microarray Experiments
- Computing Asymptotic Power and Sample Size for Case-Control Genetic Association Studies in the Presence of Phenotype and/or Genotype Misclassification Errors