Abstract
AbstractOxford Nanopore (ONT) is a leading long-read technology which has been revolutionizing transcriptome analysis through its capacity to sequence the majority of transcripts from end-to-end. This has greatly increased our ability to study the diversity of transcription mechanisms such as transcription initiation, termination, and alternative splicing. However, ONT still suffers from high error rates which have thus far limited its scope to reference-based analyses. When a reference is not available or is not a viable option due to reference-bias, error correction is a crucial step towards the reconstruction of the sequenced transcripts and downstream sequence analysis of transcripts. In this paper, we present a novel computational method to error correct ONT cDNA sequencing data, called isONcorrect. IsONcorrect is able to jointly use all isoforms from a gene during error correction, thereby allowing it to correct reads at low sequencing depths. We are able to obtain a median accuracy of 98.9-99.6%, demonstrating the feasibility of applying cost-effective cDNA full transcript length sequencing for reference-free transcriptome analysis.
Publisher
Cold Spring Harbor Laboratory
Reference43 articles.
1. Transcript Profiling Using Long-Read Sequencing Technologies;Methods in Molecular Biology,2018
2. Nanopore Long-Read RNAseq Reveals Widespread Transcriptional Variation among the Surface Receptors of Individual B Cells;Nature Communications,2017
3. Realizing the potential of full-length transcriptome sequencing
4. Byrne, Ashley , Megan A. Supple , Roger Volden , Kristin L. Laidre , Beth Shapiro , and Christopher Vollmers . 2019. “Depletion of Hemoglobin Transcripts and Long Read Sequencing Improves the Transcriptome Annotation of the Polar Bear (Ursus Maritimus).” https://doi.org/10.1101/527978.
5. Chikhi, Rayan , Jan Holub , and Paul Medvedev . 2019. “Data Structures to Represent Sets of K-Long DNA Sequences.” arXiv.
Cited by
2 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献