Author:
Magalhães Wagner C.S.,Araujo Nathalia M.,Leal Thiago P.,Araujo Gilderlanio S.,Viriato Paula J.S.,Kehdy Fernanda S.,Costa Gustavo N.,Barreto Mauricio L.,Horta Bernardo L.,Lima-Costa Maria Fernanda,Pereira Alexandre C.,Tarazona-Santos Eduardo,Rodrigues Maíra R.,
Abstract
EPIGEN-Brazil is one of the largest Latin American initiatives at the interface of human genomics, public health, and computational biology. Here, we present two resources to address two challenges to the global dissemination of precision medicine and the development of the bioinformatics know-how to support it. To address the underrepresentation of non-European individuals in human genome diversity studies, we present the EPIGEN-5M+1KGP imputation panel—the fusion of the public 1000 Genomes Project (1KGP) Phase 3 imputation panel with haplotypes derived from the EPIGEN-5M data set (a product of the genotyping of 4.3 million SNPs in 265 admixed individuals from the EPIGEN-Brazil Initiative). When we imputed a target SNPs data set (6487 admixed individuals genotyped for 2.2 million SNPs from the EPIGEN-Brazil project) with the EPIGEN-5M+1KGP panel, we gained 140,452 more SNPs in total than when using the 1KGP Phase 3 panel alone and 788,873 additional high confidence SNPs (info score ≥ 0.8). Thus, the major effect of the inclusion of the EPIGEN-5M data set in this new imputation panel is not only to gain more SNPs but also to improve the quality of imputation. To address the lack of transparency and reproducibility of bioinformatics protocols, we present a conceptual Scientific Workflow in the form of a website that models the scientific process (by including publications, flowcharts, masterscripts, documents, and bioinformatics protocols), making it accessible and interactive. Its applicability is shown in the context of the development of our EPIGEN-5M+1KGP imputation panel. The Scientific Workflow also serves as a repository of bioinformatics resources.
Funder
Brazilian Ministry of Health
Department of Science and Technology from the Secretaria de Ciência, Tecnologia e Insumos Estratégicos
Financiadora de Estudos e Projetos
Brazilian Ministry of Education
Brazilian National Research Council
Minas Gerais State Agency for Support of Research
Minas Gerais Network of Population Genomics and Precision Medicine
Publisher
Cold Spring Harbor Laboratory
Subject
Genetics(clinical),Genetics
Cited by
19 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献