Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Fungal genomes: suffering with functional annotation errors

View through CrossRef
Abstract Background The genome sequence data of more than 65985 species are publicly available as of October 2021 within the National Center for Biotechnology Information (NCBI) database alone and additional genome sequences are available in other databases and also continue to accumulate at a rapid pace. However, an error-free functional annotation of these genome is essential for the research communities to fully utilize these data in an optimum and efficient manner. Results An analysis of proteome sequence data of 689 fungal species (7.15 million protein sequences) was conducted to identify the presence of functional annotation errors. Proteins associated with calcium signaling events, including calcium dependent protein kinases (CDPKs), calmodulins (CaM), calmodulin-like (CML) proteins, WRKY transcription factors, selenoproteins, and proteins associated with the terpene biosynthesis pathway, were targeted in the analysis. Gene associated with CDPKs and selenoproteins are known to be absent in fungal genomes. Our analysis, however, revealed the presence of proteins that were functionally annotated as CDPK proteins. However, InterproScan analysis indicated that none of the protein sequences annotated as “calcium dependent protein kinase” were found to encode calcium binding EF-hands at the regulatory domain. Similarly, none of a protein sequences annotated as a “selenocysteine” were found to contain a Sec (U) amino acid. Proteins annotated as CaM and CMLs also had significant discrepancies. CaM proteins should contain four calcium binding EF-hands, however, a range of 2–4 calcium binding EF-hands were present in the fungal proteins that were annotated as CaM proteins. Similarly, CMLs should possess four calcium binding EF-hands, but some of the CML annotated fungal proteins possessed either three or four calcium binding EF-hands. WRKY transcription factors are characterized by the presence of a WRKY domain and are confined to the plant kingdom. Several fungal proteins, however, were annotated as WRKY transcription factors, even though they did not contain a WRKY domain. Conclusion The presence of functional annotation errors in fungal genome and proteome databases is of considerable concern and needs to be addressed in a timely manner.
Title: Fungal genomes: suffering with functional annotation errors
Description:
Abstract Background The genome sequence data of more than 65985 species are publicly available as of October 2021 within the National Center for Biotechnology Information (NCBI) database alone and additional genome sequences are available in other databases and also continue to accumulate at a rapid pace.
However, an error-free functional annotation of these genome is essential for the research communities to fully utilize these data in an optimum and efficient manner.
Results An analysis of proteome sequence data of 689 fungal species (7.
15 million protein sequences) was conducted to identify the presence of functional annotation errors.
Proteins associated with calcium signaling events, including calcium dependent protein kinases (CDPKs), calmodulins (CaM), calmodulin-like (CML) proteins, WRKY transcription factors, selenoproteins, and proteins associated with the terpene biosynthesis pathway, were targeted in the analysis.
Gene associated with CDPKs and selenoproteins are known to be absent in fungal genomes.
Our analysis, however, revealed the presence of proteins that were functionally annotated as CDPK proteins.
However, InterproScan analysis indicated that none of the protein sequences annotated as “calcium dependent protein kinase” were found to encode calcium binding EF-hands at the regulatory domain.
Similarly, none of a protein sequences annotated as a “selenocysteine” were found to contain a Sec (U) amino acid.
Proteins annotated as CaM and CMLs also had significant discrepancies.
CaM proteins should contain four calcium binding EF-hands, however, a range of 2–4 calcium binding EF-hands were present in the fungal proteins that were annotated as CaM proteins.
Similarly, CMLs should possess four calcium binding EF-hands, but some of the CML annotated fungal proteins possessed either three or four calcium binding EF-hands.
WRKY transcription factors are characterized by the presence of a WRKY domain and are confined to the plant kingdom.
Several fungal proteins, however, were annotated as WRKY transcription factors, even though they did not contain a WRKY domain.
Conclusion The presence of functional annotation errors in fungal genome and proteome databases is of considerable concern and needs to be addressed in a timely manner.

Related Results

NICU Medication Errors: Describing the Cause and Nature of Medication Errors in a NICU in Qatar
NICU Medication Errors: Describing the Cause and Nature of Medication Errors in a NICU in Qatar
IntroductionA medication error can be defined as “any error occurring in the medication use process” and focuses on problems with the delivery of medication to a patient [1]. Medic...
Distant suffering : on mediated suffering and health care
Distant suffering : on mediated suffering and health care
<p dir="ltr">This thesis examines distant suffering, or when suffering of others who are far away is mediated or conveyed. Distant suffering often refers to when popular medi...
Distant suffering : on mediated suffering and health care
Distant suffering : on mediated suffering and health care
<p dir="ltr">This thesis examines distant suffering, or when suffering of others who are far away is mediated or conveyed. Distant suffering often refers to when popular medi...
High-quality functional genome annotation through an intercampus competition initiative
High-quality functional genome annotation through an intercampus competition initiative
Ensuring high-quality functional annotations in newly sequenced genomes has become a fundamental problem in next-generation sequencing genomics. This problem takes additional relev...
Benchmarking Hayai-Annotation Plants: A Re-evaluation Using Standard Evaluation Metrics
Benchmarking Hayai-Annotation Plants: A Re-evaluation Using Standard Evaluation Metrics
Abstract The rapid growth of next-generation sequencing (NGS) technology has led to a surge in the determination of whole genome sequences in pla...
Inferring fungal growth rates from optical density data
Inferring fungal growth rates from optical density data
AbstractQuantifying fungal growth underpins our ability to effectively treat severe fungal infections. Current methods quantify fungal growth rates from time-course morphology-spec...
An extensible genome annotation workbench based on the Galaxy Platform
An extensible genome annotation workbench based on the Galaxy Platform
Introduction Falling costs of genetic sequencing have allowed sequencing and annotation of the genomes of non-model organism. In annotating non-mod...
Mining sequence annotation databanks for association patterns
Mining sequence annotation databanks for association patterns
Abstract Motivation: Millions of protein sequences currently being deposited to sequence databanks will never be annotated manually. Similarity-based annotation gene...

Back to Top