Author:
Saitou Marie,Dahl Andy,Wang Qingbo,Liu Xuanyao
Abstract
AbstractGenome-wide association studies (GWAS) are overwhelmingly biased toward European ancestries. Nearly all existing studies agree that transferring genetic predictions from European ancestries to other populations results in a substantial loss of accuracy. This is commonly referred to as low portability of polygenic risk scores (PRS) and is one of the most important barriers to the ethical clinical deployment of PRS. Yet, it remains unclear how much various genetic factors, such as linkage disequilibrium (LD) differences, allele frequency differences or causal effect differences, contribute to low PRS portability. In this study, we used gene expression levels in lymphoblastoid cell lines (LCLs) as a simplified model of complex traits with minimal environmental variation, in order to understand how much each genetic factor contributes to PRS portability from European to African populations. We found thatcis-genetic effects on gene expression are highly similar between European and African individuals (). This stands in stark contrast to the very low estimates ofcis-genetic correlation between Europeans and Africans in previous studies, which we demonstrate are artifacts of statistical bias. We showed that portability decreases with increasing LD differences in thecis-regions. We also found that allele frequency differences of causal variants have a striking impact on PRS portability. For example, PRS portability is reduced by more than 32% when the causalcis-variant is common (minor allele frequency, MAF > 5%) in European samples (training population) but is rarer (MAF < 5%) in African samples (prediction population). While large allele frequency differences can decrease PRS portability through increasing LD differences, we also show that causal allele frequency can significantly impact portability independently of LD. This observation suggests that improving statistical fine-mapping alone does not overcome the loss of portability caused by causal allele frequency differences. Lastly, we also found that causal allele frequency is the main genetic factor underlying differential gene expression levels across ancestries. We conclude that causal genetic effects are highly similar in Europeans and Africans, and low PRS portability is primarily due to allele frequency differences.
Publisher
Cold Spring Harbor Laboratory
Cited by
12 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献