arrow
返回

Cracking the genetic code with neural networks

delete2023-04-06
delete2
delete
OA
AI
M
Marc Joiret *
M
Marine Leclercq
G
Gaspard Lambrechts
F
Francesca Rapino
P
Pierre Close
G
Gilles Louppe
L
Liesbet Geris
DOI:10.3389/frai.2023.1128153delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
The genetic code is textbook scientific knowledge that was soundly established without resorting to Artificial Intelligence (AI). The goal of our study was to check whether a neural network could re-discover, on its own, the mapping links between codons and amino acids and build the complete deciphering dictionary upon presentation of transcripts proteins data training pairs. We compared different Deep Learning neural network architectures and estimated quantitatively the size of the required human transcriptomic training set to achieve the best possible accuracy in the codon-to-amino-acid mapping. We also investigated the effect of a codon embedding layer assessing the semantic similarity between codons on the rate of increase of the training accuracy. We further investigated the benefit of quantifying and using the unbalanced representations of amino acids within real human proteins for a faster deciphering of rare amino acids codons. Deep neural networks require huge amount of data to train them. Deciphering the genetic code by a neural network is no exception. A test accuracy of 100% and the unequivocal deciphering of rare codons such as the tryptophan codon or the stop codons require a training dataset of the order of 4-22 millions cumulated pairs of codons with their associated amino acids presented to the neural network over around 7-40 training epochs, depending on the architecture and settings. We confirm that the wide generic capacities and modularity of deep neural networks allow them to be customized easily to learn the deciphering task of the genetic code efficiently.
Keyword:
Artificial Intelligence
genetic code deciphering
codon usage
codon embedding
deep neural network
data efficiency
natural language processing
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

F
Frontiers in Artificial Intelligence
IF:
4.7
论文数:
2.4K
被引数:
4.4K

机构

U
University of Liege
学者数:
1.7W
论文数: 1.4W
被引数: 2.1W
引用论文

引用论文

Coliquefaction of Lignite and Corn Stalk in Ethanol–Water Mixed Solvent with Addition of Formic Acid and Iron Ore Catalyst
err2018-02-06
err0
PREAI
errHengfu Shui; Hanren Jiao; Fang He; Tao Shui; Xiaoling Wang; Hua Zhou; Chunxiu Pan; Zhicai Wang; Zhiping Lei; Shibiao Ren; Shigang Kang; Charles Chunbao Xu
err分享
err收藏
Combining hypothesis- and data-driven neuroscience modeling in FAIR workflows
err2022-07-06
err15
errOAAI
errEriksson, Olivia; Bhalla, Upinder Singh; Blackwell, Kim T.; Crook, Sharon M.; Keller, Daniel; Kramer, Andrei; Linne, Marja-Leena; Saudargiene, Ausra; Wade, Rebecca C.; Kotaleski, Jeanette Hellgren
err分享
err收藏
Chemotherapy for malignant gliomas of the brain: a review of ten-years experience
err1990-03-01
err0
PREAI
errP. Paoletti; G. Butti; R. Knerich; P. Gaetani; R. Assietti
err分享
err收藏
err分享
err收藏
学者 查看更多内容