Javascript must be enabled to continue!
Detecting haplotype-specific transcript variation in long reads with FLAIR2
View through CrossRef
Abstract
Background
RNA-Seq has brought forth significant discoveries regarding aberrations in RNA processing, implicating these RNA variants in a variety of diseases. Aberrant splicing and single nucleotide variants in RNA have been demonstrated to alter transcript stability, localization, and function. In particular, the upregulation of ADAR, an enzyme which mediates adenosine-to-inosine editing, has been previously linked to an increase in the invasiveness of lung ADC cells and associated with splicing regulation. Despite the functional importance of studying splicing and SNVs, short read RNA-Seq has limited the community’s ability to interrogate both forms of RNA variation simultaneously.
Results
We employed long-read technology to obtain full-length transcript sequences, elucidating cis-effects of variants on splicing changes at a single molecule level. We have developed a computational workflow that augments FLAIR, a tool that calls isoform models expressed in long-read data, to integrate RNA variant calls with the associated isoforms that bear them. We generated nanopore data with high sequence accuracy of H1975 lung adenocarcinoma cells with and without knockdown of
ADAR
. We applied our workflow to identify key inosine-isoform associations to help clarify the prominence of ADAR in tumorigenesis.
Conclusions
Ultimately, we find that a long-read approach provides valuable insight toward characterizing the relationship between RNA variants and splicing patterns.
Highlights
FLAIR2 has improved transcript isoform detection and incorporates sequence variants for haplotype-specific transcript detection.
In addition to haplotype-specific variant detection, it identifies transcript-specific RNA editing
Able to identify haplotype-specific transcript isoform bias in expression
Long-read sequencing identifies hyperedited transcripts that are missed from short-read sequencing methods for a more comprehensive identification of ADAR targets
Title: Detecting haplotype-specific transcript variation in long reads with FLAIR2
Description:
Abstract
Background
RNA-Seq has brought forth significant discoveries regarding aberrations in RNA processing, implicating these RNA variants in a variety of diseases.
Aberrant splicing and single nucleotide variants in RNA have been demonstrated to alter transcript stability, localization, and function.
In particular, the upregulation of ADAR, an enzyme which mediates adenosine-to-inosine editing, has been previously linked to an increase in the invasiveness of lung ADC cells and associated with splicing regulation.
Despite the functional importance of studying splicing and SNVs, short read RNA-Seq has limited the community’s ability to interrogate both forms of RNA variation simultaneously.
Results
We employed long-read technology to obtain full-length transcript sequences, elucidating cis-effects of variants on splicing changes at a single molecule level.
We have developed a computational workflow that augments FLAIR, a tool that calls isoform models expressed in long-read data, to integrate RNA variant calls with the associated isoforms that bear them.
We generated nanopore data with high sequence accuracy of H1975 lung adenocarcinoma cells with and without knockdown of
ADAR
.
We applied our workflow to identify key inosine-isoform associations to help clarify the prominence of ADAR in tumorigenesis.
Conclusions
Ultimately, we find that a long-read approach provides valuable insight toward characterizing the relationship between RNA variants and splicing patterns.
Highlights
FLAIR2 has improved transcript isoform detection and incorporates sequence variants for haplotype-specific transcript detection.
In addition to haplotype-specific variant detection, it identifies transcript-specific RNA editing
Able to identify haplotype-specific transcript isoform bias in expression
Long-read sequencing identifies hyperedited transcripts that are missed from short-read sequencing methods for a more comprehensive identification of ADAR targets.
Related Results
Importance of transcript variants in transcriptome analyses
Importance of transcript variants in transcriptome analyses
Abstract
RNA sequencing (RNA-Seq) has become a widely adopted genome-wide technique for investigating gene expression patterns. However, conventi...
Haplotype Matching with GBWT for Pangenome Graphs
Haplotype Matching with GBWT for Pangenome Graphs
Traditionally, variations from a linear reference genome were used to represent large sets of haplotypes compactly. In the linear reference genome based paradigm, the positional Bu...
WEPP: Phylogenetic Placement Achieves Near-Haplotype Resolution in Wastewater-Based Epidemiology
WEPP: Phylogenetic Placement Achieves Near-Haplotype Resolution in Wastewater-Based Epidemiology
Wastewater carries the full spectrum of pathogens and their variants infecting a population, making it a powerful resource for public health surveillance. During the SARS-CoV-2 pan...
SPEARS: Standard Performance Evaluation of Ancestral haplotype Reconstruction through Simulation
SPEARS: Standard Performance Evaluation of Ancestral haplotype Reconstruction through Simulation
Abstract
Motivation
Ancestral haplotype maps provide useful information about genomic variation and insights into biologi...
GraphK-LR: Enhancing Long-read Metagenomic Binning with Read-overlap Graphs Across Microbial Kingdoms
GraphK-LR: Enhancing Long-read Metagenomic Binning with Read-overlap Graphs Across Microbial Kingdoms
Abstract
Background: Metagenomics, the study of genetic material from environmental samples, relies on binning - the process of grouping DNA sequences from the same organis...
Long-read error correction: a survey and qualitative comparison
Long-read error correction: a survey and qualitative comparison
Abstract
Third generation sequencing technologies Pacific Biosciences and Oxford Nanopore Technologies were respectively made available in 2011 and 2014. In contras...
Hybrid correction of highly noisy Oxford Nanopore long reads using a variable-order de Bruijn graph
Hybrid correction of highly noisy Oxford Nanopore long reads using a variable-order de Bruijn graph
Abstract
Motivation
The recent rise of long read sequencing technologies such as Pacific Biosciences and Oxford Nanopore allows...
SPEARS: Standard Performance Evaluation of Ancestral Reconstruction through Simulation
SPEARS: Standard Performance Evaluation of Ancestral Reconstruction through Simulation
Motivation
Ancestral haplotype maps provide useful information about genomic variation and biological processes. Reconstructing the descendent haplotype structure o...

