Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Detecting haplotype-specific transcript variation in long reads with FLAIR2

View through CrossRef
Abstract Background RNA-Seq has brought forth significant discoveries regarding aberrations in RNA processing, implicating these RNA variants in a variety of diseases. Aberrant splicing and single nucleotide variants in RNA have been demonstrated to alter transcript stability, localization, and function. In particular, the upregulation of ADAR, an enzyme which mediates adenosine-to-inosine editing, has been previously linked to an increase in the invasiveness of lung ADC cells and associated with splicing regulation. Despite the functional importance of studying splicing and SNVs, short read RNA-Seq has limited the community’s ability to interrogate both forms of RNA variation simultaneously. Results We employed long-read technology to obtain full-length transcript sequences, elucidating cis-effects of variants on splicing changes at a single molecule level. We have developed a computational workflow that augments FLAIR, a tool that calls isoform models expressed in long-read data, to integrate RNA variant calls with the associated isoforms that bear them. We generated nanopore data with high sequence accuracy of H1975 lung adenocarcinoma cells with and without knockdown of ADAR . We applied our workflow to identify key inosine-isoform associations to help clarify the prominence of ADAR in tumorigenesis. Conclusions Ultimately, we find that a long-read approach provides valuable insight toward characterizing the relationship between RNA variants and splicing patterns. Highlights FLAIR2 has improved transcript isoform detection and incorporates sequence variants for haplotype-specific transcript detection. In addition to haplotype-specific variant detection, it identifies transcript-specific RNA editing Able to identify haplotype-specific transcript isoform bias in expression Long-read sequencing identifies hyperedited transcripts that are missed from short-read sequencing methods for a more comprehensive identification of ADAR targets
Title: Detecting haplotype-specific transcript variation in long reads with FLAIR2
Description:
Abstract Background RNA-Seq has brought forth significant discoveries regarding aberrations in RNA processing, implicating these RNA variants in a variety of diseases.
Aberrant splicing and single nucleotide variants in RNA have been demonstrated to alter transcript stability, localization, and function.
In particular, the upregulation of ADAR, an enzyme which mediates adenosine-to-inosine editing, has been previously linked to an increase in the invasiveness of lung ADC cells and associated with splicing regulation.
Despite the functional importance of studying splicing and SNVs, short read RNA-Seq has limited the community’s ability to interrogate both forms of RNA variation simultaneously.
Results We employed long-read technology to obtain full-length transcript sequences, elucidating cis-effects of variants on splicing changes at a single molecule level.
We have developed a computational workflow that augments FLAIR, a tool that calls isoform models expressed in long-read data, to integrate RNA variant calls with the associated isoforms that bear them.
We generated nanopore data with high sequence accuracy of H1975 lung adenocarcinoma cells with and without knockdown of ADAR .
We applied our workflow to identify key inosine-isoform associations to help clarify the prominence of ADAR in tumorigenesis.
Conclusions Ultimately, we find that a long-read approach provides valuable insight toward characterizing the relationship between RNA variants and splicing patterns.
Highlights FLAIR2 has improved transcript isoform detection and incorporates sequence variants for haplotype-specific transcript detection.
In addition to haplotype-specific variant detection, it identifies transcript-specific RNA editing Able to identify haplotype-specific transcript isoform bias in expression Long-read sequencing identifies hyperedited transcripts that are missed from short-read sequencing methods for a more comprehensive identification of ADAR targets.

Related Results

Importance of transcript variants in transcriptome analyses
Importance of transcript variants in transcriptome analyses
Abstract RNA sequencing (RNA-Seq) has become a widely adopted genome-wide technique for investigating gene expression patterns. However, conventi...
Haplotype Matching with GBWT for Pangenome Graphs
Haplotype Matching with GBWT for Pangenome Graphs
Traditionally, variations from a linear reference genome were used to represent large sets of haplotypes compactly. In the linear reference genome based paradigm, the positional Bu...
WEPP: Phylogenetic Placement Achieves Near-Haplotype Resolution in Wastewater-Based Epidemiology
WEPP: Phylogenetic Placement Achieves Near-Haplotype Resolution in Wastewater-Based Epidemiology
Wastewater carries the full spectrum of pathogens and their variants infecting a population, making it a powerful resource for public health surveillance. During the SARS-CoV-2 pan...
GraphK-LR: Enhancing Long-read Metagenomic Binning with Read-overlap Graphs Across Microbial Kingdoms
GraphK-LR: Enhancing Long-read Metagenomic Binning with Read-overlap Graphs Across Microbial Kingdoms
Abstract Background: Metagenomics, the study of genetic material from environmental samples, relies on binning - the process of grouping DNA sequences from the same organis...
Long-read error correction: a survey and qualitative comparison
Long-read error correction: a survey and qualitative comparison
Abstract Third generation sequencing technologies Pacific Biosciences and Oxford Nanopore Technologies were respectively made available in 2011 and 2014. In contras...
Hybrid correction of highly noisy Oxford Nanopore long reads using a variable-order de Bruijn graph
Hybrid correction of highly noisy Oxford Nanopore long reads using a variable-order de Bruijn graph
Abstract Motivation The recent rise of long read sequencing technologies such as Pacific Biosciences and Oxford Nanopore allows...
SPEARS: Standard Performance Evaluation of Ancestral Reconstruction through Simulation
SPEARS: Standard Performance Evaluation of Ancestral Reconstruction through Simulation
Motivation Ancestral haplotype maps provide useful information about genomic variation and biological processes. Reconstructing the descendent haplotype structure o...

Back to Top