Back

PA-FGRS is a novel estimator of pedigree-based genetic liability that complements genotype-based inferences into the genetic architecture of major depressive disorder

Dybdahl Krebs, M.; Georgii Hellberg, K.-L.; Lundberg, M.; Appadurai, V.; Ohlsson, H.; Pedersen, E. M.; Steinbach, J.; Matthews, J.; LaBianca, S.; Calle Sanchez, X.; Meijsen, J.; iPSYCH Study Consortium, ; Ingasson, A.; Buil Demur, A.; Vilhjalmsson, B. J.; Flint, J.; Bacanu, S.-A.; Cai, N.; Dahl, A. W.; Zaitlen, N.; Werge, T.; Kendler, K. S.; Schork, A.

2023-06-29 genetic and genomic medicine
10.1101/2023.06.23.23291611 medRxiv
Show abstract

Large biobank samples provide an opportunity to integrate broad phenotyping, familial records, and molecular genetics data to study complex traits and diseases. We introduce Pearson-Aitken Family Genetic Risk Scores (PA-FGRS), a new method for estimating disease liability from patterns of diagnoses in extended, age-censored genealogical records. We then apply the method to study a paradigmatic complex disorder, Major Depressive Disorder (MDD), using the iPSYCH2015 case-cohort study of 30,949 MDD cases, 39,655 random population controls, and more than 2 million relatives. We show that combining PA-FGRS liabilities estimated from family records with molecular genotypes of probands improves the three lines of inquiry. Incorporating PA-FGRS liabilities improves classification of MDD over and above polygenic scores, identifies robust genetic contributions to clinical heterogeneity in MDD associated with comorbidity, recurrence, and severity, and can improve the power of genome-wide association studies (GWAS). Our method is flexible and easy to use and our study approaches are generalizable to other data sets and other complex traits and diseases.

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.