返回
摘要
En 中文
Next-generation sequencing technologies revolutionized the ways in which genetic information is obtained and have opened the door for many essential applications in biomedical sciences. Hundreds of gigabytes of data are being produced, and all applications are affected by the errors in the data. Many programs have been designed to correct these errors, most of them targeting the data produced by the dominant technology of Illumina. We present a thorough comparison of these programs. Both HiSeq and MiSeq types of Illumina data are analyzed, and correcting performance is evaluated as the gain in depth and breadth of coverage, as given by correct reads and k-mers. Time and memory requirements, scalability and parallelism are considered as well. Practical guidelines are provided for the effective use of these tools. We also evaluate the efficiency of the current state-of-the-art programs for correcting Illumina data and provide research directions for further improvement.
Keyword:
DNA sequencing
Illumina data
error correction
coverage depth
coverage breadth
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.7
论文数:
5.6K
被引数:
2.7W
机构
引用论文
De novo assembly of human genomes with massively parallel short read sequencing通过大规模并行短读段测序对人类基因组进行从头组装
GENOME RESEARCH
IF5.5
Functional analysis and quantitative determination of the expression profile of human parvovirus B19
Virology
IF0
Velvet: Algorithms for de novo short read assembly using de Bruijn graphsVelvet: 使用de Bruijn图进行从头短读组装的算法
GENOME RESEARCH
IF5.5

