The goal of association mapping is to identify genetic variants that predict disease, and as the field of human genetics matures, the number of successful association studies is increasing. Many such studies have shown that for many diseases, risk is explained by a reasonably large number of variants that each explains a very small amount of disease risk. This is prompting the use of genetic risk scores in building predictive models, where information across several variants is combined for predictive modeling. In the current study, we compare the performance of four previously proposed genetic risk score methods and present a new method for constructing genetic risk score that incorporates explained variance information. The methods compared include: a simple count Genetic Risk Score, an odds ratio weighted Genetic Risk Score, a direct logistic regression Genetic Risk Score, a polygenic Genetic Risk Score, and the new explained variance weighted Genetic Risk Score. We compare the methods using a wide range of simulations in two steps, with a range of the number of deleterious single nucleotide polymorphisms (SNPs) explaining disease risk, genetic modes, baseline penetrances, sample sizes, relative risks (RR) and minor allele frequencies (MAF). Several measures of model performance were compared including overall power, C-statistic and Akaike’s Information Criterion. Our results show the relative performance of methods differs significantly, with the new explained variance weighted GRS (EV-GRS) generally performing favorably to the other methods.
Contents
- Article
-
Requires Authentication UnlicensedA New Explained-Variance Based Genetic Risk Score for Predictive Modeling of Disease RiskLicensedSeptember 25, 2012
-
Requires Authentication UnlicensedHessian Calculation for Phylogenetic Likelihood based on the Pruning Algorithm and its ApplicationsLicensedSeptember 25, 2012
-
Requires Authentication UnlicensedCluster-Localized Sparse Logistic Regression for SNP DataLicensedAugust 14, 2012
-
Requires Authentication UnlicensedHow to analyze many contingency tables simultaneously in genetic association studiesLicensedJuly 27, 2012
-
Requires Authentication UnlicensedIncorporating the Empirical Null Hypothesis into the Benjamini-Hochberg ProcedureLicensedJuly 26, 2012
-
Requires Authentication UnlicensedEstimating the Number of One-step Beneficial MutationsLicensedJuly 19, 2012
-
Requires Authentication UnlicensedTesting clonality of three and more tumors using their loss of heterozygosity profilesLicensedJuly 13, 2012
-
Requires Authentication UnlicensedCorrection for Founder Effects in Host-Viral Association Studies via Principal ComponentsLicensedJuly 12, 2012
-
Requires Authentication UnlicensedA Non-Homogeneous Dynamic Bayesian Network with Sequentially Coupled Interaction Parameters for Applications in Systems and Synthetic BiologyLicensedJuly 12, 2012
-
Requires Authentication UnlicensedAn Integrated Hierarchical Bayesian Model for Multivariate eQTL MappingLicensedJuly 12, 2012
-
Requires Authentication UnlicensedA Novel and Fast Normalization Method for High-Density ArraysLicensedJuly 12, 2012
-
Requires Authentication UnlicensedPerformance of MAX Test and Degree of Dominance Index in Predicting the Mode of InheritanceLicensedJune 27, 2012
-
Requires Authentication UnlicensedA Bayesian autoregressive three-state hidden Markov model for identifying switching monotonic regimes in Microarray time course dataLicensedJune 27, 2012
-
Requires Authentication UnlicensedQTL Mapping Using a Memetic Algorithm with Modifications of BIC as Fitness FunctionLicensedMay 18, 2012
-
Requires Authentication UnlicensedComputing Posterior Probabilities for Score-based Alignments Using ppALIGNLicensedMay 16, 2012