Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Kernelized approach enables explainable gene prioritizations for complex traits

View through CrossRef
Abstract Genome-wide association studies (GWAS) have identified numerous variant-trait associations; yet, assigning effector genes to GWAS loci remains challenging. Similarity-based machine-learning methods, such as PoPS, prioritize effector genes from shared functional profiles among trait-relevant genes. These models assign a prioritization score for each gene and nominate a single effector gene within a GWAS locus. However, the scores provide limited insight into why a gene was prioritized or whether the nomination is biologically plausible. To address this gap, we introduce Kernelized Polygenic Priority Score, K-PoPS, a kernelized reformulation of PoPS that enables gene-centric explanations by decomposing each prediction into contributions from training genes. For each prioritized gene, K-PoPS reports top contributor genes and an anchor score that quantifies support from a user-defined set of trait-relevant genes. Across 38 Pan-UK Biobank traits, the full-feature OLS implementation underlying K-PoPS improved closest-gene enrichment relative to default PoPS for 26 of 37 evaluable traits. Across 25 traits with curated anchor sets, predictions supported by anchor scores were more enriched for closest-gene proxies than unsupported predictions. When applying to blood level apolipoprotein B, K-PoPS nominated SCARB1 over UBC gene, and further provided convincing explanations that support this prediction. Using explanation evidence, K-PoPS identified multiple plausible effector genes within a dilated cardiomyopathy locus, contrary to the parsimonious assumption. In summary, K-PoPS provides a post hoc framework for examining and interpreting GWAS effector-gene nominations.
Title: Kernelized approach enables explainable gene prioritizations for complex traits
Description:
Abstract Genome-wide association studies (GWAS) have identified numerous variant-trait associations; yet, assigning effector genes to GWAS loci remains challenging.
Similarity-based machine-learning methods, such as PoPS, prioritize effector genes from shared functional profiles among trait-relevant genes.
These models assign a prioritization score for each gene and nominate a single effector gene within a GWAS locus.
However, the scores provide limited insight into why a gene was prioritized or whether the nomination is biologically plausible.
To address this gap, we introduce Kernelized Polygenic Priority Score, K-PoPS, a kernelized reformulation of PoPS that enables gene-centric explanations by decomposing each prediction into contributions from training genes.
For each prioritized gene, K-PoPS reports top contributor genes and an anchor score that quantifies support from a user-defined set of trait-relevant genes.
Across 38 Pan-UK Biobank traits, the full-feature OLS implementation underlying K-PoPS improved closest-gene enrichment relative to default PoPS for 26 of 37 evaluable traits.
Across 25 traits with curated anchor sets, predictions supported by anchor scores were more enriched for closest-gene proxies than unsupported predictions.
When applying to blood level apolipoprotein B, K-PoPS nominated SCARB1 over UBC gene, and further provided convincing explanations that support this prediction.
Using explanation evidence, K-PoPS identified multiple plausible effector genes within a dilated cardiomyopathy locus, contrary to the parsimonious assumption.
In summary, K-PoPS provides a post hoc framework for examining and interpreting GWAS effector-gene nominations.

Related Results

Gene dosage architecture across complex traits and common diseases 
Gene dosage architecture across complex traits and common diseases 
Introduction : Copy Number Variants (CNVs) are genomic deletions or duplications larger than 1000 base pairs (the fundamental unit of double-stranded nucleic acid...
Expression and polymorphism of genes in gallstones
Expression and polymorphism of genes in gallstones
ABSTRACT Through the method of clinical case control study, to explore the expression and genetic polymorphism of KLF14 gene (rs4731702 and rs972283) and SR-B1 gene...
Human-centric and Semantics-based Explainable Event Detection: A Survey
Human-centric and Semantics-based Explainable Event Detection: A Survey
Abstract In recent years, there has been a surge in interest in artificial intelligent systems that can provide human-centric explanations for decisions or predictions. No ...
Hotspot prioritizations show sensitivity to data type
Hotspot prioritizations show sensitivity to data type
ABSTRACT Prioritizing regions for conservation is essential for effectively allocating limited conservation resources. One of the most common approaches to prioriti...
XA4C: eXplainable representation learning via Autoencoders revealing Critical genes
XA4C: eXplainable representation learning via Autoencoders revealing Critical genes
ABSTRACT Machine Learning models have been frequently used in transcriptome analyses. Particularly, Representation Learning (RL), e.g., autoencod...
Genetic study of reproductive, dairy and growth traits in Guzerá cattle
Genetic study of reproductive, dairy and growth traits in Guzerá cattle
The Guzerá breed is an important Brazilian genetic resource and has been widely used as a pure breed and in crossbreeding strategies to produce animals adapted to tropical climatic...
Personality Traits of Minority Arab Teachers in the Arab Educational System in Israel
Personality Traits of Minority Arab Teachers in the Arab Educational System in Israel
 AbstractThe present research examined the personality traits prevalent among Arab teachers as a minority in the Arab educational system in Israel.Research on personality traits ha...

Back to Top