Accuracy and reproducibility of somatic point mutation calling in clinical-type targeted sequencing data-Reference-Cited by-同舟云学术

Accuracy and reproducibility of somatic point mutation calling in clinical-type targeted sequencing data

Published:2020-10-15 Issue:1 Volume:13 Page:
ISSN:1755-8794
Container-title:BMC Medical Genomics
language:en
Short-container-title:BMC Med Genomics

Author:

Karimnezhad Ali,Palidwor Gareth A.,Thavorn Kednapa,Stewart David J.,Campbell Pearl A.,Lo Bryan,Perkins Theodore J.^ORCID

Abstract

Abstract Background Treating cancer depends in part on identifying the mutations driving each patient’s disease. Many clinical laboratories are adopting high-throughput sequencing for assaying patients’ tumours, applying targeted panels to formalin-fixed paraffin-embedded tumour tissues to detect clinically-relevant mutations. While there have been some benchmarking and best practices studies of this scenario, much variant calling work focuses on whole-genome or whole-exome studies, with fresh or fresh-frozen tissue. Thus, definitive guidance on best choices for sequencing platforms, sequencing strategies, and variant calling for clinical variant detection is still being developed. Methods Because ground truth for clinical specimens is rarely known, we used the well-characterized Coriell cell lines GM12878 and GM12877 to generate data. We prepared samples to mimic as closely as possible clinical biopsies, including formalin fixation and paraffin embedding. We evaluated two well-known targeted sequencing panels, Illumina’s TruSight 170 hybrid-capture panel and the amplification-based Oncomine Focus panel. Sequencing was performed on an Illumina NextSeq500 and an Ion Torrent PGM respectively. We performed multiple replicates of each assay, to test reproducibility. Finally, we applied four different freely-available somatic single-nucleotide variant (SNV) callers to the data, along with the vendor-recommended callers for each sequencing platform. Results We did not observe major differences in variant calling success within the regions that each panel covers, but there were substantial differences between callers. All had high sensitivity for true SNVs, but numerous and non-overlapping false positives. Overriding certain default parameters to make them consistent between callers substantially reduced discrepancies, but still resulted in high false positive rates. Intersecting results from multiple replicates or from different variant callers eliminated most false positives, while maintaining sensitivity. Conclusions Reproducibility and accuracy of targeted clinical sequencing results depend less on sequencing platform and panel than on variability between replicates and downstream bioinformatics. Differences in variant callers’ default parameters are a greater influence on algorithm disagreement than other differences between the algorithms. Contrary to typical clinical practice, we recommend employing multiple variant calling pipelines and/or analyzing replicate samples, as this greatly decreases false positive calls.

Publisher

Springer Science and Business Media LLC

Subject

Genetics(clinical),Genetics

Link

https://link.springer.com/content/pdf/10.1186/s12920-020-00803-z.pdf

Reference37 articles.

1. Forbes SA, Bindal N, Bamford S, Cole C, Kok CY, Beare D, Jia M, Shepherd R, Leung K, Menzies A, et al. Cosmic: mining complete cancer genomes in the catalogue of somatic mutations in cancer. Nucleic Acids Res. 2010; 39(suppl_1):945–50.

2. de BONO JS, Ashworth A. Translating cancer research into targeted therapeutics. Nature. 2010; 467(7315):543–9.

3. Mok TS, Wu Y-L, Thongprasert S, Yang C-H, Chu D-T, Saijo N, Sunpaweravong P, Han B, Margono B, Ichinose Y, et al. Gefitinib or carboplatin–paclitaxel in pulmonary adenocarcinoma. N Engl J Med. 2009; 361(10):947–57.

4. Wong KM, Hudson TJ, McPherson JD. Unraveling the genetics of cancer: genome sequencing and beyond. Ann Rev Genom Hum Genet. 2011; 12:407–30.

5. Morgensztern D, Devarakonda S, Mitsudomi T, Maher C, Govindan R. Mutational events in lung cancer: Present and developing technologies. In: IASLC Thoracic Oncology (Second Edition). Philadelphia: Elsevier: 2018. p. 95–103. https://www.sciencedirect.com/science/article/pii/B9780323523578120013.

Cited by 10 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Placental somatic mutation in human stillbirth and live birth: A pilot case-control study of paired placental, fetal, and maternal whole genomes;Placenta;2024-09

2. IMPROVE: a feature model to predict neoepitope immunogenicity through broad-scale validation of T-cell recognition;Frontiers in Immunology;2024-04-03

3. Empirical Bayes single nucleotide variant-calling for next-generation sequencing data;Scientific Reports;2024-01-18

4. Multicentric pilot study to standardize clinical whole exome sequencing (WES) for cancer patients;npj Precision Oncology;2023-10-20

5. Identification of Somatic Mutations in Plasma Cell-Free DNA from Patients with Metastatic Oral Squamous Cell Carcinoma;International Journal of Molecular Sciences;2023-06-20