Quantifying the genetic separability of disease subtypes
En palabras de los autores
Genetic prediction of clinically defined disease subtypes could advance precision medicine, but polygenic risk scores (PRSs) often show limited ability to distinguish subtypes, either because the subtypes share genetic architecture or because available PRSs are inaccurate. We introduce GenSep, a statistical framework that estimates the oracle case-case AUC attainable if the true additive genetic values of both subtypes were known, and the proportion of this genetic separation recovered by subtype-specific PRSs. Across 18 subtype pairs in UK Biobank, oracle AUCs ranged from 0.605 to 1.000, whereas current PRSs achieved a median observed case-case AUC of only 0.530 (range, 0.510-0.754), corresponding to recovery of a median of 4.1% (range, 0.3-15.5%) of the genetic separation variance available to a perfect predictor. Oracle AUCs were generally consistent across UK Biobank, All of Us, FinnGen and the Million Veteran Program, and across European, African and Admixed American ancestry groups. FinnGen-trained PRSs, despite 2.6-to 3.9-fold larger discovery samples, performed no better than internally trained PRSs, and substantially larger consortium GWAS raised case-case AUC by 0.06 on average (range, 0.003-0.089). GenSep separates subtype pairs limited by intrinsic genetic overlap from those limited by current PRS accuracy, identifying where larger studies and better predictors could improve genetic discrimination.
Apareció: jueves, 24 de septiembre. medRxiv. Preprint, todavía sin revisión por pares.