Quantitative Biology > Populations and Evolution
This paper has been withdrawn by Sanzo Miyazawa
[Submitted on 30 Dec 2016 (v1), last revised 2 Apr 2017 (this version, v2)]
Title:Selection originating from protein foldability: I. A new method to estimate selection temperature
No PDF available, click to view other formatsAbstract:The probability distribution of sequences with maximum entropy that satisfies a given amino acid composition at each site and a given pairwise amino acid frequency at each site pair is a Boltzmann distribution with $\exp(-\psi_N)$, where the total interaction $\psi_N$ is represented as the sum of one body and pairwise interactions. A protein folding theory based on the random energy model (REM) indicates that the equilibrium ensemble of natural protein sequences is a canonical ensemble characterized by $\exp(-\Delta G_{ND}/k_B T_s)$ or by $\exp(- G_{N}/k_B T_s)$ if an amino acid composition is kept constant, meaning $\psi_N = \Delta G_{ND}/k_B T_s +$ constant, where $\Delta G_{ND} \equiv G_N - G_D$, $G_N$ and $G_D$ are the native and denatured free energies, and $T_s$ is the effective temperature of natural selection. Here, we examine interaction changes ($\Delta \psi_N$) due to single nucleotide nonsynonymous mutations, and have found that the variance of their $\Delta \psi_N$ over all sites hardly depends on the $\psi_N$ of each homologous sequence, indicating that the variance of $\Delta G_N (= k_B T_s \Delta \psi_N)$ is nearly constant irrespective of protein families. As a result, $T_s$ is estimated from the ratio of the variance of $\Delta \psi_N$ to that of a reference protein, which is determined by a direct comparison between $\Delta\Delta \psi_{ND} (\simeq \Delta \psi_N)$ and experimental $\Delta\Delta G_{ND}$. Based on the REM, glass transition temperature $T_g$ and $\Delta G_{ND}$ are estimated from $T_s$ and experimental melting temperatures ($T_m$) for 14 protein domains. The estimates of $\Delta G_{ND}$ agree well with their experimental values for 5 proteins, and those of $T_s$ and $T_g$ are all within a reasonable range. This method is coarse-grained but much simpler in estimating $T_s$, $T_g$ and $\Delta\Delta G_{ND}$ than previous methods.
Submission history
From: Sanzo Miyazawa [view email][v1] Fri, 30 Dec 2016 04:11:09 UTC (319 KB)
[v2] Sun, 2 Apr 2017 04:54:32 UTC (1 KB) (withdrawn)
Current browse context:
q-bio.PE
References & Citations
Bibliographic and Citation Tools
Bibliographic Explorer (What is the Explorer?)
Litmaps (What is Litmaps?)
scite Smart Citations (What are Smart Citations?)
Code, Data and Media Associated with this Article
CatalyzeX Code Finder for Papers (What is CatalyzeX?)
DagsHub (What is DagsHub?)
Gotit.pub (What is GotitPub?)
Papers with Code (What is Papers with Code?)
ScienceCast (What is ScienceCast?)
Demos
Recommenders and Search Tools
Influence Flower (What are Influence Flowers?)
Connected Papers (What is Connected Papers?)
CORE Recommender (What is CORE?)
arXivLabs: experimental projects with community collaborators
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.
Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.
Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.