返回
Detecting Selective Sweeps from Pooled Next-Generation Sequencing Samples
DOI:10.1093/molbev/mss090.png)
摘要
En 中文
Due to its cost effectiveness, next-generation sequencing of pools of individuals (Pool-Seq) is becoming a popular strategy for characterizing variation in population samples. Because Pool-Seq provides genome-wide SNP frequency data, it is possible to use them for demographic inference and/or the identification of selective sweeps. Here, we introduce a statistical method that is designed to detect selective sweeps from pooled data by accounting for statistical challenges associated with Pool-Seq, namely sequencing errors and random sampling among chromosomes. This allows for an efficient use of the information: all base calls are included in the analysis, but the higher credibility of regions with higher coverage and base calls with better quality scores is accounted for. Computer simulations show that our method efficiently detects sweeps even at very low coverage (0.5x per chromosome). Indeed, the power of detecting sweeps is similar to what we could expect from sequences of individual chromosomes. Since the inference of selective sweeps is based on the allele frequency spectrum (AFS), we also provide a method to accurately estimate the AFS provided that the quality scores for the sequence reads are reliable. Applying our approach to Pool-Seq data from Drosophila melanogaster, we identify several selective sweep signatures on chromosome X that include some previously well-characterized sweeps like the wapl region.
Keyword:
selective sweeps
next-generation sequencing
pooled DNA
Drosophila
allele frequency spectrum
hidden Markov model
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
5.3
论文数:
8.3K
被引数:
6.6W
机构
引用论文
The Ly6/uPAR protein Bouncer is necessary and sufficient for species-specific fertilization
Science
IF0
Multilocus patterns of nucleotide variability and the demographic and selection history of Drosophila melanogaster populations
GENOME RESEARCH
IF5.5
Substantial biases in ultra-short read data sets from high-throughput DNA sequencing来自高通量DNA测序的超短读取数据集中的实质性偏差
NUCLEIC ACIDS RESEARCH
IF13.1
The Next Generation of Molecular Markers From Massively Parallel Sequencing of Pooled DNA Samples
GENETICS
IF5.1

