Javascript must be enabled to continue!
ePath: an online database towards comprehensive essential gene annotation for prokaryotes
View through CrossRef
Abstract
Experimental techniques for identification of essential genes (EGs) in prokaryotes are usually expensive, time-consuming and sometimes unrealistic. Emerging
in silico
methods provide alternative methods for EG prediction, but often possess limitations including heavy computational requirements and lack of biological explanation. Here we propose a new computational algorithm for EG prediction in prokaryotes with an online database (ePath) for quick access to the EG prediction results of over 4,000 prokaryotes (
https://www.pubapps.vcu.edu/epath/
). In ePath, gene essentiality is linked to biological functions annotated by KEGG Ortholog (KO). Two new scoring systems, namely, E_score and P_score, are proposed for each KO as the EG evaluation criteria. E_score represents appearance and essentiality of a given KO in existing experimental results of gene essentiality, while P_score denotes gene essentiality based on the principle that a gene is essential if it plays a role in genetic information processing, cell envelope maintenance or energy production. The new EG prediction algorithm shows prediction accuracy ranging from 75% to 91% based on validation from five new experimental studies on EG identification. Our overall goal with ePath is to provide a comprehensive and reliable reference for gene essentiality annotation, facilitating the study of those prokaryotes without experimentally derived gene essentiality information.
Springer Science and Business Media LLC
Title: ePath: an online database towards comprehensive essential gene annotation for prokaryotes
Description:
Abstract
Experimental techniques for identification of essential genes (EGs) in prokaryotes are usually expensive, time-consuming and sometimes unrealistic.
Emerging
in silico
methods provide alternative methods for EG prediction, but often possess limitations including heavy computational requirements and lack of biological explanation.
Here we propose a new computational algorithm for EG prediction in prokaryotes with an online database (ePath) for quick access to the EG prediction results of over 4,000 prokaryotes (
https://www.
pubapps.
vcu.
edu/epath/
).
In ePath, gene essentiality is linked to biological functions annotated by KEGG Ortholog (KO).
Two new scoring systems, namely, E_score and P_score, are proposed for each KO as the EG evaluation criteria.
E_score represents appearance and essentiality of a given KO in existing experimental results of gene essentiality, while P_score denotes gene essentiality based on the principle that a gene is essential if it plays a role in genetic information processing, cell envelope maintenance or energy production.
The new EG prediction algorithm shows prediction accuracy ranging from 75% to 91% based on validation from five new experimental studies on EG identification.
Our overall goal with ePath is to provide a comprehensive and reliable reference for gene essentiality annotation, facilitating the study of those prokaryotes without experimentally derived gene essentiality information.
Related Results
Principes et outils pour l’annotation des corpus
Principes et outils pour l’annotation des corpus
La linguistique de corpus, c’est à dire les recherches sur le langage portant sur un matériel linguistique écrit ou oral recueilli et conservé, s’est considérablement développée au...
Impact of Gene Annotation Choice on the Quantification of RNA-Seq Data
Impact of Gene Annotation Choice on the Quantification of RNA-Seq Data
Abstract
Background: RNA sequencing is currently the method of choice for genome-wide profiling of gene expression. A popular approach to quantify expression levels of gene...
Benchmarking Hayai-Annotation Plants: A Re-evaluation Using Standard Evaluation Metrics
Benchmarking Hayai-Annotation Plants: A Re-evaluation Using Standard Evaluation Metrics
Abstract
The rapid growth of next-generation sequencing (NGS) technology has led to a surge in the determination of whole genome sequences in pla...
Expression and polymorphism of genes in gallstones
Expression and polymorphism of genes in gallstones
ABSTRACT
Through the method of clinical case control study, to explore the expression and genetic polymorphism of KLF14 gene (rs4731702 and rs972283) and SR-B1 gene...
An extensible genome annotation workbench based on the Galaxy Platform
An extensible genome annotation workbench based on the Galaxy Platform
Introduction
Falling costs of genetic sequencing have allowed sequencing and annotation of the genomes of non-model organism. In annotating non-mod...
Developing an eHealth Tool to Support Patient Empowerment at Home
Developing an eHealth Tool to Support Patient Empowerment at Home
In previous research we have learned that patients with chronic or complex diseases often experience difficulties when transitioning from hospital care to self-care in their home. ...
FAMUS: A Few-Shot Learning Framework for Large-Scale Protein Annotation
FAMUS: A Few-Shot Learning Framework for Large-Scale Protein Annotation
Predicting gene function is a pivotal and challenging step in genomic and metagenomic data analysis. Current automatic annotation tools typically rely on the single most similar se...
Pinaceae show elevated rates of gene duplication and gene loss that are robust to incomplete gene annotation
Pinaceae show elevated rates of gene duplication and gene loss that are robust to incomplete gene annotation
Abstract
Gene duplications and gene losses are major determinants of genome evolution and phenotypic diversity. The frequency of gene turnover (gene gains and gene ...

