Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Can we use it? On the utility of de novo and reference-based assembly of Nanopore data for plant plastome sequencing

View through CrossRef
Abstract The chloroplast genome harbors plenty of valuable information for phylogenetic research. Illumina short-read data is generally used for de novo assembly of whole plastomes. PacBio or Oxford Nanopore long reads are additionally employed in hybrid approaches to enable assembly across the highly similar inverted repeats of a chloroplast genome. Unlike for PacBio, plastome assemblies based solely on Nanopore reads are rarely found, due to their high error rate and non-random error profile. However, the actual quality decline connected to their use has never been quantified. Furthermore, no study has employed reference-based assembly using Nanopore reads, which is common with Illumina data. Using Leucanthemum Mill. as an example, we compared the sequence quality of seven plastome assemblies of the same species, using combinations of two sequencing platforms and three analysis pipelines. In addition, we assessed the factors which might influence Nanopore assembly quality during sequence generation and bioinformatic processing. The consensus sequence derived from de novo assembly of Nanopore data had a sequence identity of 99.59% compared to Illumina short-read de novo assembly. Most of the found errors comprise indels (81.5%), and a large majority of them is part of homopolymer regions. The quality of reference-based assembly is heavily dependent upon the choice of a close-enough reference. Using a reference with 0.83% sequence divergence from the studied species, mapping of Nanopore reads results in a consensus comparable to that from Nanopore de novo assembly, and of only slightly inferior quality compared to a reference-based assembly with Illumina data (0.49% and 0.26% divergence from Illumina de novo ). For optimal assembly of Nanopore data, appropriate filtering of contaminants and chimeric sequences, as well as employing moderate read coverage, is essential. Based on these results, we conclude that Nanopore long reads are a suitable alternative to Illumina short reads in plastome phylogenomics. Only few errors remain in the finalized assembly, which can be easily masked in phylogenetic analyses without loss in analytical accuracy. The easily applicable and cost-effective technology might warrant more attention by researchers dealing with plant chloroplast genomes.
Title: Can we use it? On the utility of de novo and reference-based assembly of Nanopore data for plant plastome sequencing
Description:
Abstract The chloroplast genome harbors plenty of valuable information for phylogenetic research.
Illumina short-read data is generally used for de novo assembly of whole plastomes.
PacBio or Oxford Nanopore long reads are additionally employed in hybrid approaches to enable assembly across the highly similar inverted repeats of a chloroplast genome.
Unlike for PacBio, plastome assemblies based solely on Nanopore reads are rarely found, due to their high error rate and non-random error profile.
However, the actual quality decline connected to their use has never been quantified.
Furthermore, no study has employed reference-based assembly using Nanopore reads, which is common with Illumina data.
Using Leucanthemum Mill.
as an example, we compared the sequence quality of seven plastome assemblies of the same species, using combinations of two sequencing platforms and three analysis pipelines.
In addition, we assessed the factors which might influence Nanopore assembly quality during sequence generation and bioinformatic processing.
The consensus sequence derived from de novo assembly of Nanopore data had a sequence identity of 99.
59% compared to Illumina short-read de novo assembly.
Most of the found errors comprise indels (81.
5%), and a large majority of them is part of homopolymer regions.
The quality of reference-based assembly is heavily dependent upon the choice of a close-enough reference.
Using a reference with 0.
83% sequence divergence from the studied species, mapping of Nanopore reads results in a consensus comparable to that from Nanopore de novo assembly, and of only slightly inferior quality compared to a reference-based assembly with Illumina data (0.
49% and 0.
26% divergence from Illumina de novo ).
For optimal assembly of Nanopore data, appropriate filtering of contaminants and chimeric sequences, as well as employing moderate read coverage, is essential.
Based on these results, we conclude that Nanopore long reads are a suitable alternative to Illumina short reads in plastome phylogenomics.
Only few errors remain in the finalized assembly, which can be easily masked in phylogenetic analyses without loss in analytical accuracy.
The easily applicable and cost-effective technology might warrant more attention by researchers dealing with plant chloroplast genomes.

Related Results

Plastomes of the green algae Hydrodictyon reticulatum and Pediastrum duplex (Sphaeropleales, Chlorophyceae)
Plastomes of the green algae Hydrodictyon reticulatum and Pediastrum duplex (Sphaeropleales, Chlorophyceae)
Background Comparative studies of chloroplast genomes (plastomes) across the Chlorophyceae are revealing dynamic patterns of size variation, gene content, and genome...
Monitoring airborne pathogens by nanopore sequencing
Monitoring airborne pathogens by nanopore sequencing
Next generation sequencing technologies have revolutionized the field of environmental science. Widely used short-read sequencing enables accurate microbial identification but is o...
Factors Affecting the Survival of Therapy Related AML/MDS, Secondary AML and De Novo AML
Factors Affecting the Survival of Therapy Related AML/MDS, Secondary AML and De Novo AML
Abstract Introduction: Patients exposed to cytotoxic agents are at a higher risk of developing therapy related AML and MDS (tAML/tMDS), and have poor ...
Assigning Transcriptomic Subtypes to Chronic Lymphocytic Leukemia Samples Using Nanopore RNA-Sequencing and Self-Organizing Maps
Assigning Transcriptomic Subtypes to Chronic Lymphocytic Leukemia Samples Using Nanopore RNA-Sequencing and Self-Organizing Maps
Background/Objectives: Massively parallel sequencing technologies have advanced chronic lymphocytic leukemia (CLL) diagnostics and precision oncology. Illumina platforms, while off...
Nanopore pour la détection de peptides vers le séquençage de protéines à l'échelle de la molécule unique.
Nanopore pour la détection de peptides vers le séquençage de protéines à l'échelle de la molécule unique.
Les protéines sont des macromolécules indispensables au fonctionnement des cellules vivantes, pouvant fournir des informations clés pour comprendre les mécanismes biologiques e...
Quantitative detection of DNA methylation from nanopore sequencing data without raw signals
Quantitative detection of DNA methylation from nanopore sequencing data without raw signals
Abstract Background Nanopore sequencing has revolutionized the field of epigenomics by enabling direct detection of DNA m...
Microrna Regulation of Nodule Zone-Specific Gene Expression In Soybean
Microrna Regulation of Nodule Zone-Specific Gene Expression In Soybean
Nitrogen is a paramount important essential element for all living organisms. It has been found to bea crucial structural component of proteins, nucleic acids, enzymes and other ce...

Back to Top