Javascript must be enabled to continue!
Virtual ChIP-seq: predicting transcription factor binding by learning from the transcriptome
View through CrossRef
Abstract
Motivation
Identifying transcription factor binding sites is the first step in pinpointing non-coding mutations that disrupt the regulatory function of transcription factors and promote disease. ChIP-seq is the most common method for identifying binding sites, but performing it on patient samples is hampered by the amount of available biological material and the cost of the experiment. Existing methods for computational prediction of regulatory elements primarily predict binding in genomic regions with sequence similarity to known transcription factor sequence preferences. This has limited efficacy since most binding sites do not resemble known transcription factor sequence motifs, and many transcription factors are not even sequence-specific.
Results
We developed Virtual ChIP-seq, which predicts binding of individual transcription factors in new cell types using an artificial neural network that integrates ChIP-seq results from other cell types and chromatin accessibility data in the new cell type. Virtual ChIP-seq also uses learned associations between gene expression and transcription factor binding at specific genomic regions. This approach outperforms methods that predict TF binding solely based on sequence preference, pre-dicting binding for 36 transcription factors (Matthews correlation coefficient > 0.3).
Availability
The datasets we used for training and validation are available at
https://virchip.hoffmanlab.org
. We have deposited in Zenodo the current version of our software (
http://doi.org/10.5281/zenodo.1066928
), datasets (
http://doi.org/10.5281/zenodo.823297
), predictions for 36 transcription factors on Roadmap Epigenomics cell types (
http://doi.org/10.5281/zenodo.1455759
), and predictions in Cistrome as well as ENCODE-DREAM
in vivo
TF Binding Site Prediction Challenge (
http://doi.org/10.5281/zenodo.1209308
).
Title: Virtual ChIP-seq: predicting transcription factor binding by learning from the transcriptome
Description:
Abstract
Motivation
Identifying transcription factor binding sites is the first step in pinpointing non-coding mutations that disrupt the regulatory function of transcription factors and promote disease.
ChIP-seq is the most common method for identifying binding sites, but performing it on patient samples is hampered by the amount of available biological material and the cost of the experiment.
Existing methods for computational prediction of regulatory elements primarily predict binding in genomic regions with sequence similarity to known transcription factor sequence preferences.
This has limited efficacy since most binding sites do not resemble known transcription factor sequence motifs, and many transcription factors are not even sequence-specific.
Results
We developed Virtual ChIP-seq, which predicts binding of individual transcription factors in new cell types using an artificial neural network that integrates ChIP-seq results from other cell types and chromatin accessibility data in the new cell type.
Virtual ChIP-seq also uses learned associations between gene expression and transcription factor binding at specific genomic regions.
This approach outperforms methods that predict TF binding solely based on sequence preference, pre-dicting binding for 36 transcription factors (Matthews correlation coefficient > 0.
3).
Availability
The datasets we used for training and validation are available at
https://virchip.
hoffmanlab.
org
.
We have deposited in Zenodo the current version of our software (
http://doi.
org/10.
5281/zenodo.
1066928
), datasets (
http://doi.
org/10.
5281/zenodo.
823297
), predictions for 36 transcription factors on Roadmap Epigenomics cell types (
http://doi.
org/10.
5281/zenodo.
1455759
), and predictions in Cistrome as well as ENCODE-DREAM
in vivo
TF Binding Site Prediction Challenge (
http://doi.
org/10.
5281/zenodo.
1209308
).
Related Results
KEDUDUKAN AHLI BAHASA DALAM PEMBUKTIAN PERKARA PENCEMARAN NAMA BAIK (STUDI PUTUSAN NOMOR: 47/PID.SUS/2019/PN. MGT)
KEDUDUKAN AHLI BAHASA DALAM PEMBUKTIAN PERKARA PENCEMARAN NAMA BAIK (STUDI PUTUSAN NOMOR: 47/PID.SUS/2019/PN. MGT)
<em><span id="page3R_mcid52" class="markedContent"><span style="left: calc(var(--scale-factor)*125.30px); top: calc(var(--scale-factor)*539.11px); font-size: calc(va...
MARS-seq2.0: an experimental and analytical pipeline for indexed sorting combined with single-cell RNA sequencing v1
MARS-seq2.0: an experimental and analytical pipeline for indexed sorting combined with single-cell RNA sequencing v1
Human tissues comprise trillions of cells that populate a complex space of molecular phenotypes and functions and that vary in abundance by 4–9 orders of magnitude. Relying solely ...
A plug and play microfluidic platform for standardized sensitive low-input chromatin immunoprecipitation
A plug and play microfluidic platform for standardized sensitive low-input chromatin immunoprecipitation
Epigenetic profiling by chromatin immunoprecipitation followed by sequencing (ChIP-seq) has become a powerful tool for genome-wide identification of regulatory elements, for defini...
A plug and play microfluidic platform for standardized sensitive low-input Chromatin Immunoprecipitation
A plug and play microfluidic platform for standardized sensitive low-input Chromatin Immunoprecipitation
Abstract
Epigenetic profiling by ChIP-Seq has become a powerful tool for genome-wide identification of regulatory elements, for defining transcriptional regulatory ...
Comparing genome-wide chromatin profiles using ChIP-chip or ChIP-seq
Comparing genome-wide chromatin profiles using ChIP-chip or ChIP-seq
AbstractMotivation: ChIP-chip and ChIP-seq technologies provide genome-wide measurements of various types of chromatin marks at an unprecedented resolution. With ChIP samples colle...
Unmeasured human transcription factor ChIP-seq data shape functional genomics and demand strategic prioritization
Unmeasured human transcription factor ChIP-seq data shape functional genomics and demand strategic prioritization
Abstract
Transcription factor (TF) chromatin immunoprecipitation followed by sequencing (ChIP-seq) is essential for identifying genome-wide TF-binding sites (TFBSs),...
CREATING LEARNING MEDIA IN TEACHING ENGLISH AT SMP MUHAMMADIYAH 2 PAGELARAN ACADEMIC YEAR 2020/2021
CREATING LEARNING MEDIA IN TEACHING ENGLISH AT SMP MUHAMMADIYAH 2 PAGELARAN ACADEMIC YEAR 2020/2021
The pandemic Covid-19 currently demands teachers to be able to use technology in teaching and learning process. But in reality there are still many teachers who have not been able ...
Benchmarking Algorithms for Gene Set Scoring of Single-cell ATAC-seq Data
Benchmarking Algorithms for Gene Set Scoring of Single-cell ATAC-seq Data
AbstractGene set scoring (GSS) has been routinely conducted for gene expression analysis of bulk or single-cell RNA-seq data, which helps to decipher single-cell heterogeneity and ...

