Expanded explorations into the optimization of an energy function for protein design.
IEEE/ACM Trans. Comput. Biology Bioinform.(2013)
摘要
Nature possesses a secret formula for the energy as a function of the structure of a protein. In protein design, approximations are made to both the structural representation of the molecule and to the form of the energy equation, such that the existence of a general energy function for proteins is by no means guaranteed. Here, we present new insights toward the application of machine learning to the problem of finding a general energy function for protein design. Machine learning requires the definition of an objective function, which carries with it the implied definition of success in protein design. We explored four functions, consisting of two functional forms, each with two criteria for success. Optimization was carried out by a Monte Carlo search through the space of all variable parameters. Cross-validation of the optimized energy function against a test set gave significantly different results depending on the choice of objective function, pointing to relative correctness of the built-in assumptions. Novel energy cross terms correct for the observed nonadditivity of energy terms and an imbalance in the distribution of predicted amino acids. This paper expands on the work presented at the 2012 ACM-BCB.
更多查看译文
关键词
optimisation,energy function,physics,rotamers,learning (artificial intelligence),novel energy cross term,monte carlo methods,objective function,energy equation,proteins,chemistry,biology and genetics,expanded explorations,molecular biophysics,general energy function,optimization,correlation,protein design,implied definition,dead-end elimination,structural representation,energy term,optimized energy function,machine learning,amino acids,monte carlo search,bioinformatics,learning artificial intelligence,hydrogen,linear programming
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络