arrow
返回

A supervised Bayesian method for time (re)annotation of transcriptomics data

delete2025-12-01
delete0
PRE
AI
E
Elio Nushi
F
François P. Douillard
K
Katja Selby
B
Benjamin A. Blount
O
Oliver Pennington
N
Nigel P. Minton
M
Miia Lindström
A
Antti Honkela *
DOI:10.1093/nargab/lqaf203delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Transcriptomics experiments are often conducted to capture changes in gene expression over time. However, time annotations may be missing, imprecise, or not reflect the same physiological state of the bacterial culture between different experiments. Assigning accurate time points to these experiments using a reference time course is therefore crucial for identifying differentially expressed genes, and understanding gene regulatory networks for elucidating the studied organism's physiology and life cycle. This important task, which could enhance the biological interpretation of the transcriptomics experiments, has not been previously addressed. In this work, we propose a novel method to solve the challenge of realigning transcriptomics experiments based on a reference time course. Our method is based on a Bayesian approach that uses Gaussian process regression modeling. We show a use case of applying our method for assigning time annotations in legacy microarray samples of the bacterium Clostridium botulinum, which were solely annotated based on the growth phase at the time when the culture aliquots were sampled, utilizing recently collected RNA-Seq time series data comprising multiple replicates as a reference. The method significantly improved the description of the growth phases of the microarray data compared to the original annotations by clearly delineating the microarray samples belonging to different growth phases, as demonstrated by principal component analysis. Consequently, a larger number of differentially expressed genes was detected when comparing experiments belonging to successive growth phases. We compare this innovative approach with a baseline method that uses k-nearest neighbor algorithm and show that our method offers a higher resolution in the description of the data by exposing smaller time changes between samples. We also test the performance of the method on sparse RNA-Seq time series (i.e. sampled every second hour). All the predictions for the samples were within a 30-min margin of their true time.
Keyword:
GAUSSIAN-PROCESSES
IDENTIFICATION

期刊

N
NAR Genomics and Bioinformatics
IF:
2.8
论文数:
223
被引数:
0

机构

U
university of helsinki
学者数:
4.1W
论文数: 3.6W
被引数: 51
U
uk research & innovation (ukri)
学者数:
2.7W
论文数: 2.3W
被引数: 32
U
university of nottingham
学者数:
3.8K
论文数: 1.8K
被引数: 0
学者 查看更多机构
引用论文

引用论文

The dynamics and regulators of cell fate decisions are revealed by pseudotemporal ordering of single cells单细胞的假时间顺序揭示了细胞命运决定的动力学和调节剂
err2014-03-23
err4.3K
errOAAI
errTrapnell, Cole; Cacchiarelli, Davide; Grimsby, Jonna; Pokharel, Prapti; Li, Shuqiang; Morse, Michael; Lennon, Niall J.; Livak, Kenneth J.; Mikkelsen, Tarjei S.; Rinn, John L.
err分享
err收藏
err分享
err收藏
Detecting time periods of differential gene expression using Gaussian processes: an application to endothelial cells exposed to radiotherapy dose fraction
err2014-10-28
err0
errOAAI
errMarkus Heinonen; Olivier Guipaud; Fabien Milliat; Valérie Buard; Béatrice Micheau; Georges Tarlet; Marc Benderitter; Farida Zehraoui; Florence d’Alché-Buc
err分享
err收藏
Model-based method for transcription factor target identification with limited data
err2010-04-12
err84
errOAAI
errHonkela, Antti; Girardot, Charles; Gustafson, E. Hilary; Liu, Ya-Hsin; Furlong, Eileen E. M.; Lawrence, Neil D.; Rattray, Magnus
err分享
err收藏
err分享
err收藏
Estimating replicate time shifts using Gaussian process regression
err2010-02-09
err0
errOAAI
errQiang Liu; Kevin K. Lin; Bogi Andersen; Padhraic Smyth; Alexander Ihler
err分享
err收藏
学者 查看更多内容