返回
Genotyping Polyploids from Messy Sequencing Data
DOI:10.1534/genetics.118.301468.png)
摘要
En 中文
Detecting and quantifying the differences in individual genomes (i.e., genotyping), plays a fundamental role in most modern bioinformatics pipelines. Many scientists now use reduced representation next-generation sequencing (NGS) approaches for genotyping. Genotyping diploid individuals using NGS is a well-studied field, and similar methods for polyploid individuals are just emerging. However, there are many aspects of NGS data, particularly in polyploids, that remain unexplored by most methods. Our contributions in this paper are fourfold: (i) We draw attention to, and then model, common aspects of NGS data: sequencing error, allelic bias, overdispersion, and outlying observations. (ii) Many datasets feature related individuals, and so we use the structure of Mendelian segregation to build an empirical Bayes approach for genotyping polyploid individuals. (iii) We develop novel models to account for preferential pairing of chromosomes, and harness these for genotyping. (iv) We derive oracle genotyping error rates that may be used for read depth suggestions. We assess the accuracy of our method in simulations, and apply it to a dataset of hexaploid sweet potato (Ipomoea batatas). An R package implementing our method is available at https://cran.r-project.org/package=updog.
Keyword:
GBS
RAD-Seq
sequencing
hierarchical modeling
read-mapping bias
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
5.1
论文数:
8.1K
被引数:
3.6W
机构
引用论文
Segregation models for disomic, tetrasomic and intermediate inheritance in tetraploids: A general procedure applied to Rorippa (Yellow cress) microsatellite data
GENETICS
IF5.1
The Genome Analysis Toolkit: A MapReduce framework for analyzing next-generation DNA sequencing data基因组分析工具包: 用于分析下一代DNA测序数据的MapReduce框架
GENOME RESEARCH
IF5.5
Low-coverage sequencing: Implications for design of complex trait association studies
GENOME RESEARCH
IF5.5


