返回
Context-aware transcript quantification from long-read RNA-seq data with Bambu
DOI:10.1038/s41592-023-01908-w.png)
摘要
En 中文
Most approaches to transcript quantification rely on fixed reference annotations; however, the transcriptome is dynamic and depending on the context, such static annotations contain inactive isoforms for some genes, whereas they are incomplete for others. Here we present Bambu, a method that performs machine-learning-based transcript discovery to enable quantification specific to the context of interest using long-read RNA-sequencing. To identify novel transcripts, Bambu estimates the novel discovery rate, which replaces arbitrary per-sample thresholds with a single, interpretable, precision-calibrated parameter. Bambu retains the full-length and unique read counts, enabling accurate quantification in presence of inactive isoforms. Compared to existing methods for transcript discovery, Bambu achieves greater precision without sacrificing sensitivity. We show that context-aware annotations improve quantification for both novel and known transcripts. We apply Bambu to quantify isoforms from repetitive HERVH-LTR7 retrotransposons in human embryonic stem cells, demonstrating the ability for context-specific transcript expression analysis. Leveraging long-read RNA-seq data and machine learning, Bambu facilitates accurate transcript discovery and quantification.
期刊
IF:
32.1
论文数:
7.2K
被引数:
12.7W
机构
引用论文
Full-length transcript characterization of SF3B1 mutation in chronic lymphocytic leukemia reveals downregulation of retained introns
NATURE COMMUNICATIONS
IF15.7
Transcriptome assembly from long-read RNA-seq alignments with StringTie2用StringTie2从长读rna-seq比对进行转录组组装
GENOME BIOLOGY
IF9.4
Performance of neural network basecalling tools for Oxford Nanopore sequencing用于牛津纳米孔测序的神经网络基础调用工具的性能
GENOME BIOLOGY
IF9.4

