We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

q-bio.GN

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Quantitative Biology > Genomics

Title: Nanopore Sequencing Technology and Tools for Genome Assembly: Computational Analysis of the Current State, Bottlenecks and Future Directions

Abstract: Nanopore sequencing technology has the potential to render other sequencing technologies obsolete with its ability to generate long reads and provide portability. However, high error rates of the technology pose a challenge while generating accurate genome assemblies. The tools used for nanopore sequence analysis are of critical importance as they should overcome the high error rates of the technology. Our goal in this work is to comprehensively analyze current publicly available tools for nanopore sequence analysis to understand their advantages, disadvantages, and performance bottlenecks. It is important to understand where the current tools do not perform well to develop better tools. To this end, we 1) analyze the multiple steps and the associated tools in the genome assembly pipeline using nanopore sequence data, and 2) provide guidelines for determining the appropriate tools for each step. We analyze various combinations of different tools and expose the tradeoffs between accuracy, performance, memory usage and scalability. We conclude that our observations can guide researchers and practitioners in making conscious and effective choices for each step of the genome assembly pipeline using nanopore sequence data. Also, with the help of bottlenecks we have found, developers can improve the current tools or build new ones that are both accurate and fast, in order to overcome the high error rates of the nanopore sequencing technology.
Comments: To appear in Briefings in Bioinformatics (BIB), 2018
Subjects: Genomics (q-bio.GN)
Journal reference: Briefings in Bioinformatics. 2019 Jul 19;20(4):1542-1559
DOI: 10.1093/bib/bby017
Cite as: arXiv:1711.08774 [q-bio.GN]
  (or arXiv:1711.08774v4 [q-bio.GN] for this version)

Submission history

From: Damla Senol Cali [view email]
[v1] Thu, 23 Nov 2017 17:03:02 GMT (494kb)
[v2] Sun, 18 Feb 2018 03:23:45 GMT (496kb)
[v3] Sat, 3 Mar 2018 22:12:12 GMT (496kb)
[v4] Tue, 6 Mar 2018 03:20:56 GMT (496kb)

Link back to: arXiv, form interface, contact.