Imputation Performance in Latin American Populations: Improving Rare Variants Representation With the Inclusion of Native American Genomes

Author:

Jiménez-Kaufmann Andrés,Chong Amanda Y.,Cortés Adrián,Quinto-Cortés Consuelo D.,Fernandez-Valverde Selene L.,Ferreyra-Reyes Leticia,Cruz-Hervert Luis Pablo,Medina-Muñoz Santiago G.,Sohail Mashaal,Palma-Martinez María J.,Delgado-Sánchez Gudalupe,Mongua-Rodríguez Norma,Mentzer Alexander J.,Hill Adrian V. S.,Moreno-Macías Hortensia,Huerta-Chagoya Alicia,Aguilar-Salinas Carlos A.,Torres Michael,Kim Hie Lim,Kalsi Namrata,Schuster Stephan C.,Tusié-Luna Teresa,Del-Vecchyo Diego Ortega,García-García Lourdes,Moreno-Estrada Andrés

Abstract

Current Genome-Wide Association Studies (GWAS) rely on genotype imputation to increase statistical power, improve fine-mapping of association signals, and facilitate meta-analyses. Due to the complex demographic history of Latin America and the lack of balanced representation of Native American genomes in current imputation panels, the discovery of locally relevant disease variants is likely to be missed, limiting the scope and impact of biomedical research in these populations. Therefore, the necessity of better diversity representation in genomic databases is a scientific imperative. Here, we expand the 1,000 Genomes reference panel (1KGP) with 134 Native American genomes (1KGP + NAT) to assess imputation performance in Latin American individuals of mixed ancestry. Our panel increased the number of SNPs above the GWAS quality threshold, thus improving statistical power for association studies in the region. It also increased imputation accuracy, particularly in low-frequency variants segregating in Native American ancestry tracts. The improvement is subtle but consistent across countries and proportional to the number of genomes added from local source populations. To project the potential improvement with a higher number of reference genomes, we performed simulations and found that at least 3,000 Native American genomes are needed to equal the imputation performance of variants in European ancestry tracts. This reflects the concerning imbalance of diversity in current references and highlights the contribution of our work to reducing it while complementing efforts to improve global equity in genomic research.

Funder

Newton Fund

Consejo Nacional de Ciencia y Tecnología

Publisher

Frontiers Media SA

Subject

Genetics (clinical),Genetics,Molecular Medicine

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3