Javascript must be enabled to continue!
Newly identified genomic sequences establish benchmarks for proposed taxonomic classification of jingmenviruses
View through CrossRef
Abstract
Jingmenviruses are a group of viruses related to orthoflaviviruses characterized by a segmented genome and multipartite organisation that have been detected worldwide in a wide range of hosts. As next-generation sequencing has become more affordable, increasing numbers of large-scale metagenomics studies have been published, alongside raw sequencing data. With the growing number of new jingmenvirus sequences identified in metagenomics data, it can be difficult to assess whether a new sequence is associated to a new virus species or to a strain of an existing species. In order to propose clear classification criteria for this group, we assembled a large jingmenvirus sequence database, from published and newly assembled sequences. Indeed, we screened data from studies that did not search for or report jingmenvirus sequences, looking for new strains of known jingmenvirus species. We then performed multiple sequence alignments and used the inferred percentage identity values to determine demarcation criteria based on the distribution of evolutionary distances upon pairwise comparisons. We report the identification of almost 60 libraries containing jingmenvirus sequences, in a wide range of sample types and geographical locations. Using these data and published jingmenvirus sequences, we have determined that to classify jingmenvirus sequences into virus species, at least four segments are required, on which eight cut-off values in percentage identity (nucleotide and amino acid) are used for demarcation. The ratification of this proposal would enhance consistency in virus taxonomy and provide a standardized framework for comparative genomics studies of jingmenviruses, a group that is yet under-characterised.
Importance
There are currently no guidelines to classify jingmenvirus sequences which means sequences are associated to species with no ratified rationale, which can result in missed opportunities at better describing jingmenvirus genomics. In this study we propose criteria to provide a parsimonious framework for classifying jingmenvirus species based on current knowledge, and allow us to propose re-classification of some sequences and to identify outlier sequences which we recommend should be further characterised
in vitro
.
Title: Newly identified genomic sequences establish benchmarks for proposed taxonomic classification of jingmenviruses
Description:
Abstract
Jingmenviruses are a group of viruses related to orthoflaviviruses characterized by a segmented genome and multipartite organisation that have been detected worldwide in a wide range of hosts.
As next-generation sequencing has become more affordable, increasing numbers of large-scale metagenomics studies have been published, alongside raw sequencing data.
With the growing number of new jingmenvirus sequences identified in metagenomics data, it can be difficult to assess whether a new sequence is associated to a new virus species or to a strain of an existing species.
In order to propose clear classification criteria for this group, we assembled a large jingmenvirus sequence database, from published and newly assembled sequences.
Indeed, we screened data from studies that did not search for or report jingmenvirus sequences, looking for new strains of known jingmenvirus species.
We then performed multiple sequence alignments and used the inferred percentage identity values to determine demarcation criteria based on the distribution of evolutionary distances upon pairwise comparisons.
We report the identification of almost 60 libraries containing jingmenvirus sequences, in a wide range of sample types and geographical locations.
Using these data and published jingmenvirus sequences, we have determined that to classify jingmenvirus sequences into virus species, at least four segments are required, on which eight cut-off values in percentage identity (nucleotide and amino acid) are used for demarcation.
The ratification of this proposal would enhance consistency in virus taxonomy and provide a standardized framework for comparative genomics studies of jingmenviruses, a group that is yet under-characterised.
Importance
There are currently no guidelines to classify jingmenvirus sequences which means sequences are associated to species with no ratified rationale, which can result in missed opportunities at better describing jingmenvirus genomics.
In this study we propose criteria to provide a parsimonious framework for classifying jingmenvirus species based on current knowledge, and allow us to propose re-classification of some sequences and to identify outlier sequences which we recommend should be further characterised
in vitro
.
Related Results
Jingmenviruses: Ubiquitous, understudied, segmented flavi-like viruses
Jingmenviruses: Ubiquitous, understudied, segmented flavi-like viruses
Jingmenviruses are a group of viruses identified recently, in 2014, and currently classified by the International Committee on Taxonomy of Viruses as unclassified Flaviviridae. The...
The fluid genomic organisation of jingmenviruses
The fluid genomic organisation of jingmenviruses
Abstract
Jingmenviruses are a distinct group of flavi-like viruses characterized by a genome consisting of four to five segments. Here, we report the discovery of t...
Phylogenetic Classification of Feline Immunodeficiency Virus
Phylogenetic Classification of Feline Immunodeficiency Virus
Background: The feline immunodeficiency virus (FIV) is responsible for a retroviral disease that affects domestic and wild cats worldwide, causing Feline Acquired Immunodeficiency ...
Partial RdRp sequences offer a robust method for Coronavirus subgenus classification
Partial RdRp sequences offer a robust method for Coronavirus subgenus classification
Abstract
The recent reclassification of the
Riboviria
, and the introduction of multiple new taxonomic catego...
Genomic benchmarks: a collection of datasets for genomic sequence classification
Genomic benchmarks: a collection of datasets for genomic sequence classification
Abstract
Background
Recently, deep neural networks have been successfully applied in many biological fields. In 2020, a d...
Use cases for Taxonomic Name Services
Use cases for Taxonomic Name Services
Catalogue of Life Plus (CoL+) was started in 2017 as a collaborative project between the Catalogue of Life (COL), the Global Biodiversity Information Facility (GBIF), Naturalis Bio...
Effective Use of Benchmarking: The Context of the Centre for Preparatory Studies in Oman
Effective Use of Benchmarking: The Context of the Centre for Preparatory Studies in Oman
This paper examines the importance and the process of developing benchmarks for the courses offered at the Centre for Preparatory Studies, Sultan Qaboos University, Oman. Benchmark...
GENERator: A Long-Context Generative Genomic Foundation Model
GENERator: A Long-Context Generative Genomic Foundation Model
Abstract
The rapid advancement of DNA sequencing has produced vast genomic datasets, yet the interpretation and rational engineering of sequence function remain fun...

