Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Probabilistic alignment leads to improved accuracy and read coverage for bisulfite sequencing data

View through CrossRef
Abstract Background DNA methylation has been linked to many important biological phenomena. Researchers have recently begun to sequence bisulfite treated DNA to determine its pattern of methylation. However, sequencing reads from bisulfite-converted DNA can vary significantly from the reference genome because of incomplete bisulfite conversion, genome variation, sequencing errors, and poor quality bases. Therefore, it is often difficult to align reads to the correct locations in the reference genome. Furthermore, bisulfite sequencing experiments have the additional complexity of having to estimate the DNA methylation levels within the sample. Results Here, we present a highly accurate probabilistic algorithm, which is an extension of the Genomic Next-generation Universal MAPper to accommodate bisulfite sequencing data (GNUMAP-bs), that addresses the computational problems associated with aligning bisulfite sequencing data to a reference genome. GNUMAP-bs integrates uncertainty from read and mapping qualities to help resolve the difference between poor quality bases and the ambiguity inherent in bisulfite conversion. We tested GNUMAP-bs and other commonly-used bisulfite alignment methods using both simulated and real bisulfite reads and found that GNUMAP-bs and other dynamic programming methods were more accurate than the more heuristic methods. Conclusions The GNUMAP-bs aligner is a highly accurate alignment approach for processing the data from bisulfite sequencing experiments. The GNUMAP-bs algorithm is freely available for download at: http://dna.cs.byu.edu/gnumap. The software runs on multiple threads and multiple processors to increase the alignment speed.
Title: Probabilistic alignment leads to improved accuracy and read coverage for bisulfite sequencing data
Description:
Abstract Background DNA methylation has been linked to many important biological phenomena.
Researchers have recently begun to sequence bisulfite treated DNA to determine its pattern of methylation.
However, sequencing reads from bisulfite-converted DNA can vary significantly from the reference genome because of incomplete bisulfite conversion, genome variation, sequencing errors, and poor quality bases.
Therefore, it is often difficult to align reads to the correct locations in the reference genome.
Furthermore, bisulfite sequencing experiments have the additional complexity of having to estimate the DNA methylation levels within the sample.
Results Here, we present a highly accurate probabilistic algorithm, which is an extension of the Genomic Next-generation Universal MAPper to accommodate bisulfite sequencing data (GNUMAP-bs), that addresses the computational problems associated with aligning bisulfite sequencing data to a reference genome.
GNUMAP-bs integrates uncertainty from read and mapping qualities to help resolve the difference between poor quality bases and the ambiguity inherent in bisulfite conversion.
We tested GNUMAP-bs and other commonly-used bisulfite alignment methods using both simulated and real bisulfite reads and found that GNUMAP-bs and other dynamic programming methods were more accurate than the more heuristic methods.
Conclusions The GNUMAP-bs aligner is a highly accurate alignment approach for processing the data from bisulfite sequencing experiments.
The GNUMAP-bs algorithm is freely available for download at: http://dna.
cs.
byu.
edu/gnumap.
The software runs on multiple threads and multiple processors to increase the alignment speed.

Related Results

BiSulfite Bolt: A bisulfite sequencing analysis platform
BiSulfite Bolt: A bisulfite sequencing analysis platform
Abstract Background Bisulfite sequencing is commonly used to measure DNA methylation. Processing bisulfite sequencing dat...
BiSulfite Bolt: A BiSulfite Sequencing Analysis Platform
BiSulfite Bolt: A BiSulfite Sequencing Analysis Platform
Abstract Background Bisulfite sequencing is commonly employed to measure DNA methylation. Processing bisulfite sequencing data ...
Inventory and pricing management in probabilistic selling
Inventory and pricing management in probabilistic selling
Context: Probabilistic selling is the strategy that the seller creates an additional probabilistic product using existing products. The exact information is unknown to customers u...
Umap and Bismap: quantifying genome and methylome mappability
Umap and Bismap: quantifying genome and methylome mappability
Abstract Motivation Short-read sequencing enables assessment of genetic and biochemical traits of individu...
The impact of environmental uncertainty on business–IT alignment: A study in Sri Lanka
The impact of environmental uncertainty on business–IT alignment: A study in Sri Lanka
<p>Despite the widely held belief that organisational performance can be enhanced through the alignment of information technology (IT) and business strategy, alignment remain...
Targeted Long-Read Bisulfite Sequencing for Promoter Methylation Analysis in Severe Preterm Birth
Targeted Long-Read Bisulfite Sequencing for Promoter Methylation Analysis in Severe Preterm Birth
Abstract DNA methylation plays a critical role in the dynamics of gene expression regulation and the development of various disorders. Whole-genome bisulfite sequen...
Plant species-specific basecaller improves actual accuracy of nanopore sequencing
Plant species-specific basecaller improves actual accuracy of nanopore sequencing
Abstract Background Long-read sequencing platforms offered by Oxford Nanopore Technologies (ONT) allow native DNA containing epigenetic modifications to be directly sequen...
Species-specific basecallers improve actual accuracy of nanopore sequencing in plants
Species-specific basecallers improve actual accuracy of nanopore sequencing in plants
Abstract Background Long-read sequencing platforms offered by Oxford Nanopore Technologies (ONT) allow native DNA containing epigenetic modification...

Back to Top