Computational codon optimization of synthetic gene for protein expression

Bevan K.S. Chung, Dong Yup Lee

Research output: Contribution to journalArticlepeer-review

70 Scopus citations

Abstract

Background: The construction of customized nucleic acid sequences allows us to have greater flexibility in gene design for recombinant protein expression. Among the various parameters considered for such DNA sequence design, individual codon usage (ICU) has been implicated as one of the most crucial factors affecting mRNA translational efficiency. However, previous works have also reported the significant influence of codon pair usage, also known as codon context (CC), on the level of protein expression.Results: In this study, we have developed novel computational procedures for evaluating the relative importance of optimizing ICU and CC for enhancing protein expression. By formulating appropriate mathematical expressions to quantify the ICU and CC fitness of a coding sequence, optimization procedures based on genetic algorithm were employed to maximize its ICU and/or CC fitness. Surprisingly, the in silico validation of the resultant optimized DNA sequences for Escherichia coli, Lactococcus lactis, Pichia pastoris and Saccharomyces cerevisiae suggests that CC is a more relevant design criterion than the commonly considered ICU.Conclusions: The proposed CC optimization framework can complement and enhance the capabilities of current gene design tools, with potential applications to heterologous protein production and even vaccine development in synthetic biotechnology.

Original languageEnglish
Article number134
JournalBMC Systems Biology
Volume6
DOIs
StatePublished - 20 Oct 2012
Externally publishedYes

Fingerprint

Dive into the research topics of 'Computational codon optimization of synthetic gene for protein expression'. Together they form a unique fingerprint.

Cite this