Loading...
Thumbnail Image
Item

Accurate Viral Population Assembly From Ultra-Deep Sequencing Data

Mangul, Serghei
Wu, Nicholas C.
Mancuso, Nicholas
Zelikovskiy, Alexander
Sun, Ren
Eskin, Eleazar
Citations
Altmetric:
Abstract

Motivation: Next-generation sequencing technologies sequence viruses with ultra-deep coverage, thus promising to revolutionize our understanding of the underlying diversity of viral populations. While the sequencing coverage is high enough that even rare viral variants are sequenced, the presence of sequencing errors makes it difficult to distinguish between rare variants and sequencing errors. Results: In this article, we present a method to overcome the limitations of sequencing technologies and assemble a diverse viral population that allows for the detection of previously undiscovered rare variants. The proposed method consists of a high-fidelity sequencing protocol and an accurate viral population assembly method, referred to as Viral Genome Assembler (VGA). The proposed protocol is able to eliminate sequencing errors by using individual barcodes attached to the sequencing fragments. Highly accurate data in combination with deep coverage allow VGA to assemble rare variants. VGA uses an expectation–maximization algorithm to estimate abundances of the assembled viral variants in the population. Results on both synthetic and real datasets show that our method is able to accurately assemble an HIV viral population and detect rare variants previously undetectable due to sequencing errors. VGA outperforms state-of-the-art methods for genome-wide viral assembly. Furthermore, our method is the first viral assembly method that scales to millions of sequencing reads.

Comments
<p>Originally Published in:</p> <p>Bioinformatics, 30 (12), i329-37. doi: <a href="http://dx.doi.org/10.1093/bioinformatics/btu295">10.1093/bioinformatics/btu295</a></p>
Description
Date
2014-06-01
Journal Title
Journal ISSN
Volume Title
Publisher
Research Projects
Organizational Units
Journal Issue
Keywords
Citation
Serghei Mangul, Nicholas C. Wu, Nicholas Mancuso, Alex Zelikovsky, Ren Sun, Eleazar Eskin; Accurate viral population assembly from ultra-deep sequencing data, Bioinformatics, Volume 30, Issue 12, 15 June 2014, Pages i329–i337, https://doi.org/10.1093/bioinformatics/btu295
Embargo Lift Date
DOI
Embedded videos