Accurate and robust genomic prediction of celiac disease using statistical learning.

Authors
Abraham, Gad 
Tye-Din, Jason A 
Bhalala, Oneil G 
Kowalczyk, Adam 
Zobel, Justin 

Loading...
Thumbnail Image
Type
Article
Change log
Abstract

Practical application of genomic-based risk stratification to clinical diagnosis is appealing yet performance varies widely depending on the disease and genomic risk score (GRS) method. Celiac disease (CD), a common immune-mediated illness, is strongly genetically determined and requires specific HLA haplotypes. HLA testing can exclude diagnosis but has low specificity, providing little information suitable for clinical risk stratification. Using six European cohorts, we provide a proof-of-concept that statistical learning approaches which simultaneously model all SNPs can generate robust and highly accurate predictive models of CD based on genome-wide SNP profiles. The high predictive capacity replicated both in cross-validation within each cohort (AUC of 0.87-0.89) and in independent replication across cohorts (AUC of 0.86-0.9), despite differences in ethnicity. The models explained 30-35% of disease variance and up to ∼43% of heritability. The GRS's utility was assessed in different clinically relevant settings. Comparable to HLA typing, the GRS can be used to identify individuals without CD with ≥99.6% negative predictive value however, unlike HLA typing, fine-scale stratification of individuals into categories of higher-risk for CD can identify those that would benefit from more invasive and costly definitive testing. The GRS is flexible and its performance can be adapted to the clinical situation by adjusting the threshold cut-off. Despite explaining a minority of disease heritability, our findings indicate a genomic risk score provides clinically relevant information to improve upon current diagnostic pathways for CD and support further studies evaluating the clinical utility of this approach in CD and other complex diseases.

Publication Date
2014-02
Online Publication Date
2014-02-13
Acceptance Date
2013-12-08
Keywords
Alleles, Biometry, Celiac Disease, Female, Genetic Predisposition to Disease, Genome, Human, Genomics, HLA Antigens, Haplotypes, Humans, Polymorphism, Single Nucleotide, Risk
Journal Title
PLoS Genet
Journal ISSN
1553-7390
1553-7404
Volume Title
10
Publisher
Public Library of Science (PLoS)