Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

A method for predicting evolved fold switchers exclusively from their sequences

View through CrossRef
Abstract Although most proteins with known structures conform to the longstanding rule-of-thumb that high levels of aligned sequence identity tend to indicate similar folds and functions, an increasing number of exceptions is emerging. In spite of having highly similar sequences, these “evolved fold switchers” (1) can adopt radically different folds with disparate biological functions. Predictive methods for identifying evolved fold switchers are desirable because some of them are associated with disease and/or can perform different functions in cells. Previously, we showed that inconsistencies between predicted and experimentally determined secondary structures can be used to predict fold switching proteins (2). The usefulness of this approach is limited, however, because it requires experimentally determined protein structures, whose magnitude is dwarfed by the number of genomic proteins. Here, we use secondary structure predictions to identify evolved fold switchers from their amino acid sequences alone. To do this, we looked for inconsistencies between the secondary structure predictions of the alternative conformations of evolved fold switchers. We used three different predictors in this study: JPred4, PSIPRED, and SPIDER3. We find that overall inconsistencies are not a significant predictor of evolved fold switchers for any of the three predictors. Inconsistencies between α-helix and β-strand predictions made by JPred4, however, can discriminate between the different conformations of evolved fold switchers with statistical significance (p < 1.7*10 −13 ). In light of this observation, we used these inconsistencies as a classifier and found that it could robustly discriminate between evolved fold switchers and evolved non-fold-switchers, as evidenced by a Matthews correlation coefficient of 0.90. These results indicate that inconsistencies between secondary structure predictions can indeed be used to identify evolved fold switchers from their genomic sequences alone. Our findings have implications for genomics, structural biology, and human health.
Title: A method for predicting evolved fold switchers exclusively from their sequences
Description:
Abstract Although most proteins with known structures conform to the longstanding rule-of-thumb that high levels of aligned sequence identity tend to indicate similar folds and functions, an increasing number of exceptions is emerging.
In spite of having highly similar sequences, these “evolved fold switchers” (1) can adopt radically different folds with disparate biological functions.
Predictive methods for identifying evolved fold switchers are desirable because some of them are associated with disease and/or can perform different functions in cells.
Previously, we showed that inconsistencies between predicted and experimentally determined secondary structures can be used to predict fold switching proteins (2).
The usefulness of this approach is limited, however, because it requires experimentally determined protein structures, whose magnitude is dwarfed by the number of genomic proteins.
Here, we use secondary structure predictions to identify evolved fold switchers from their amino acid sequences alone.
To do this, we looked for inconsistencies between the secondary structure predictions of the alternative conformations of evolved fold switchers.
We used three different predictors in this study: JPred4, PSIPRED, and SPIDER3.
We find that overall inconsistencies are not a significant predictor of evolved fold switchers for any of the three predictors.
Inconsistencies between α-helix and β-strand predictions made by JPred4, however, can discriminate between the different conformations of evolved fold switchers with statistical significance (p < 1.
7*10 −13 ).
In light of this observation, we used these inconsistencies as a classifier and found that it could robustly discriminate between evolved fold switchers and evolved non-fold-switchers, as evidenced by a Matthews correlation coefficient of 0.
90.
These results indicate that inconsistencies between secondary structure predictions can indeed be used to identify evolved fold switchers from their genomic sequences alone.
Our findings have implications for genomics, structural biology, and human health.

Related Results

Frequency of Common Chromosomal Abnormalities in Patients with Idiopathic Acquired Aplastic Anemia
Frequency of Common Chromosomal Abnormalities in Patients with Idiopathic Acquired Aplastic Anemia
Objective: To determine the frequency of common chromosomal aberrations in local population idiopathic determine the frequency of common chromosomal aberrations in local population...
Role of Organic Agriculture in Enhancing Soil Health: Implications for Physico-Chemical and Biological Properties
Role of Organic Agriculture in Enhancing Soil Health: Implications for Physico-Chemical and Biological Properties
Soil health is fundamental to sustainable agriculture and food security. Organic agricultural practices have gained increasing recognition for its capacity to improving soil health...
Sédimentologie, stratigraphie séquentielle et cyclostratigraphie du Kimméridgien du Jura suisse et du Bassin vocontien (France)
Sédimentologie, stratigraphie séquentielle et cyclostratigraphie du Kimméridgien du Jura suisse et du Bassin vocontien (France)
À l'heure actuelle, peu de travaux concernent les épaisses séries calcaires soi-disant homogènes du Kimméridgien du Jura central. Ainsi, la stratigraphie comme les facteurs (tecton...
A sequence-based approach for identifying protein fold switchers
A sequence-based approach for identifying protein fold switchers
Abstract Although most proteins conform to the classical one-structure/one-function paradigm, an increasing number of proteins with dual structures and functions ar...
Abstract 3361: Surfaceome and cancer testis antigen profiling of lung adenocarcinoma by large-scale transcriptomic analysis
Abstract 3361: Surfaceome and cancer testis antigen profiling of lung adenocarcinoma by large-scale transcriptomic analysis
Abstract Background: The emergence of improved modalities to target tumor specific and tumor associated antigens in solid tumors (e.g. antibody-drug conjugates, bi-s...
Développements méthodologiques autour de l'analyse des données de metabarcoding ADN
Développements méthodologiques autour de l'analyse des données de metabarcoding ADN
Cette thèse s'inscrit dans le cadre du traitement des données issues de séquençage haut débit, et en particulier des données produites en metabarcoding ADN. Le metabarcoding ADN co...

Back to Top