Abstract
AbstractResearch efforts of the ongoing SARS-CoV-2 pandemic have focused on viral genome sequence analysis to understand how the virus spread across the globe. Here, we assess three recently identified SARS-CoV-2 genomes in Beijing from June 2020 and attempt to determine the origin of these genomes, made available in the GISAID database. The database contains fully or partially sequenced SARS-CoV-2 samples from laboratories around the world. Including the three new samples and excluding samples with missing annotations, we analyzed 7, 643 SARS-CoV-2 genomes. Using principal component analysis computed on a similarity matrix that compares all pairs of the SARS-CoV-2 nucleotide sequences at all loci simultaneously, using the Jaccard index, we find that the newly discovered virus genomes from Beijing are in a genetic cluster that consists mostly of cases from Europe and South(east) Asia. The sequences of the new cases are most related to virus genomes from a small number of cases from China (March 2020), cases from Europe (February to early May 2020), and cases from South(east) Asia (May to June 2020). These findings could suggest that the original cases of this genetic cluster originated from China in March 2020 and were re-introduced to China by transmissions from samples from South(east) Asia between April and June 2020.
Publisher
Cold Spring Harbor Laboratory
Reference9 articles.
1. Elbe, S. and Buckland-Merrett, G. (2017). Data, disease and diplomacy: GISAID’s innovative contribution to global health. Global Challenges, 1(33–46).
2. Hahn, G. , Lee, S. , Weiss, S. , and Lange, C. (2020a). Unsupervised cluster analysis of sars-cov-2 genomes reflects its geographic progression and identifies distinct genetic subgroups of sars-cov-2 virus. Preprint at bioRxiv:2020.05.05.079061.
3. Hahn, G. , Lutz, S. , Hecker, J. , Prokopenko, D. , Cho, M. , Silverman, E. , Weiss, S. , and Lange, C. (2020b). locstra: Fast analysis of regional/global stratification in whole genome sequencing (wgs) studies. Preprint at bioRxiv:2020.03.06.981050.
4. Hahn, G. , Lutz, S. , and Lange, C. (2020c). locStra: Fast Implementation of (Local) Population Stratification Methods (v1.3). https://cran.r-project.org/web/packages/locStra/index.html.
5. Liu, R. and Nebehay, S. (2020). China sees European virus strain in Beijing, WHO says more study needed. Reuters News (2020-06-18 9:26pm). https://www.reuters.com/article/us-health-coronavirus-china-virus-data-idUSKBN23Q04L.