返回
CodSeqGen: A tool for generating synonymous coding sequences with desired GC-contents
DOI:10.1016/j.ygeno.2019.02.002.png)
摘要
En 中文
Identification of regulatory elements is essential for understanding the mechanism behind regulating gene expression. These regulatory elements-located in or near gene-bind to proteins called transcription factors to initiate the transcription process. Their occurrences are influenced by the GC-content or nucleotide composition. For generating synthetic coding sequences with pre-specified amino acid sequence and desired GC-content, there exist two stochastic methods, multinomial and maximum entropy. Both methods rely on the probability of choosing the codon synonymous for usage in regard to a specific amino acid. In spite the latter exhibited unbiased manner, the produced sequences are not exactly obeying the GC-content constraint. In this paper, we present an algorithmic solution to produce coding sequences that follow exactly a primary amino acid sequence and a desired GC-content. The proposed tool, namely CodSeqGen, depends on random selection for smaller subsets to be traversed using the backtracking approach.
Keyword:
Synonymous coding sequence
Sequence analysis
GC-content
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3
论文数:
7.2K
被引数:
1.2W
机构
引用论文
GC-Content Evolution in Bacterial Genomes: The Biased Gene Conversion Hypothesis Expands
PLOS GENETICS
IF3.7
Non-uniform slant estimation and correction for Farsi/Arabic handwritten words波斯语/阿拉伯语手写单词的非均匀倾斜估计和校正
Accounting for background nucleotide composition when measuring codon usage bias测量密码子使用偏好性时考虑背景核苷酸组成
没有更多内容

