Javascript must be enabled to continue!
Probabilistic alignment leads to improved accuracy and read coverage for bisulfite sequencing data
View through CrossRef
Abstract
Background
DNA methylation has been linked to many important biological phenomena. Researchers have recently begun to sequence bisulfite treated DNA to determine its pattern of methylation. However, sequencing reads from bisulfite-converted DNA can vary significantly from the reference genome because of incomplete bisulfite conversion, genome variation, sequencing errors, and poor quality bases. Therefore, it is often difficult to align reads to the correct locations in the reference genome. Furthermore, bisulfite sequencing experiments have the additional complexity of having to estimate the DNA methylation levels within the sample.
Results
Here, we present a highly accurate probabilistic algorithm, which is an extension of the Genomic Next-generation Universal MAPper to accommodate bisulfite sequencing data (GNUMAP-bs), that addresses the computational problems associated with aligning bisulfite sequencing data to a reference genome. GNUMAP-bs integrates uncertainty from read and mapping qualities to help resolve the difference between poor quality bases and the ambiguity inherent in bisulfite conversion. We tested GNUMAP-bs and other commonly-used bisulfite alignment methods using both simulated and real bisulfite reads and found that GNUMAP-bs and other dynamic programming methods were more accurate than the more heuristic methods.
Conclusions
The GNUMAP-bs aligner is a highly accurate alignment approach for processing the data from bisulfite sequencing experiments. The GNUMAP-bs algorithm is freely available for download at: http://dna.cs.byu.edu/gnumap. The software runs on multiple threads and multiple processors to increase the alignment speed.
Springer Science and Business Media LLC
Title: Probabilistic alignment leads to improved accuracy and read coverage for bisulfite sequencing data
Description:
Abstract
Background
DNA methylation has been linked to many important biological phenomena.
Researchers have recently begun to sequence bisulfite treated DNA to determine its pattern of methylation.
However, sequencing reads from bisulfite-converted DNA can vary significantly from the reference genome because of incomplete bisulfite conversion, genome variation, sequencing errors, and poor quality bases.
Therefore, it is often difficult to align reads to the correct locations in the reference genome.
Furthermore, bisulfite sequencing experiments have the additional complexity of having to estimate the DNA methylation levels within the sample.
Results
Here, we present a highly accurate probabilistic algorithm, which is an extension of the Genomic Next-generation Universal MAPper to accommodate bisulfite sequencing data (GNUMAP-bs), that addresses the computational problems associated with aligning bisulfite sequencing data to a reference genome.
GNUMAP-bs integrates uncertainty from read and mapping qualities to help resolve the difference between poor quality bases and the ambiguity inherent in bisulfite conversion.
We tested GNUMAP-bs and other commonly-used bisulfite alignment methods using both simulated and real bisulfite reads and found that GNUMAP-bs and other dynamic programming methods were more accurate than the more heuristic methods.
Conclusions
The GNUMAP-bs aligner is a highly accurate alignment approach for processing the data from bisulfite sequencing experiments.
The GNUMAP-bs algorithm is freely available for download at: http://dna.
cs.
byu.
edu/gnumap.
The software runs on multiple threads and multiple processors to increase the alignment speed.
Related Results
BiSulfite Bolt: A bisulfite sequencing analysis platform
BiSulfite Bolt: A bisulfite sequencing analysis platform
Abstract
Background
Bisulfite sequencing is commonly used to measure DNA methylation. Processing bisulfite sequencing dat...
BiSulfite Bolt: A BiSulfite Sequencing Analysis Platform
BiSulfite Bolt: A BiSulfite Sequencing Analysis Platform
Abstract
Background
Bisulfite sequencing is commonly employed to measure DNA methylation. Processing bisulfite sequencing data ...
Inventory and pricing management in probabilistic selling
Inventory and pricing management in probabilistic selling
Context: Probabilistic selling is the strategy that the seller creates an additional probabilistic product using existing products. The exact information is unknown to customers u...
Umap and Bismap: quantifying genome and methylome mappability
Umap and Bismap: quantifying genome and methylome mappability
Abstract
Motivation
Short-read sequencing enables assessment of genetic and biochemical traits of individu...
The impact of environmental uncertainty on business–IT alignment: A study in Sri Lanka
The impact of environmental uncertainty on business–IT alignment: A study in Sri Lanka
<p>Despite the widely held belief that organisational performance can be enhanced through the alignment of information technology (IT) and business strategy, alignment remain...
Targeted Long-Read Bisulfite Sequencing for Promoter Methylation Analysis in Severe Preterm Birth
Targeted Long-Read Bisulfite Sequencing for Promoter Methylation Analysis in Severe Preterm Birth
Abstract
DNA methylation plays a critical role in the dynamics of gene expression regulation and the development of various disorders. Whole-genome bisulfite sequen...
Plant species-specific basecaller improves actual accuracy of nanopore sequencing
Plant species-specific basecaller improves actual accuracy of nanopore sequencing
Abstract
Background
Long-read sequencing platforms offered by Oxford Nanopore Technologies (ONT) allow native DNA containing epigenetic modifications to be directly sequen...
Species-specific basecallers improve actual accuracy of nanopore sequencing in plants
Species-specific basecallers improve actual accuracy of nanopore sequencing in plants
Abstract
Background
Long-read sequencing platforms offered by Oxford Nanopore Technologies (ONT) allow native DNA containing epigenetic modification...

