Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Self-Supervised Multi-Level Generative Adversarial Network Data Imputation Algorithm

View through CrossRef
Abstract Data missing has always been a challenging problem in machine learning. The Generative Adversarial Imputation Networks (GAIN) have been shown to outperform many existing solutions. However, in GAIN, because missing values lack ground truth as supervision, it is unable to construct reconstruction loss for missing values and can only judge the reasonableness of imputed values based on reconstruction loss of non-missing values and adversarial loss. From the perspective of granular computing, data has levels, and data at different levels of granularity encapsulates different knowledge. Therefore, based on granular computing, this paper proposes a self-supervised multi-level generative adversarial network data imputation algorithm (MGAIN). Firstly, multiple levels of data are constructed using nested feature set sequences. Then, GAIN is used to impute missing values at the coarsest granularity level, and the imputation results of missing values at the coarse granularity level are used as supervision for imputing missing values at the fine granularity level, constructing reconstruction loss for missing values at the fine granularity level. Finally, based on reconstruction loss of missing values, reconstruction loss of non-missing values, and adversarial loss, data at the finer granularity level is imputed. MGAIN imputes missing values level by level from the coarse granularity level to the fine granularity level to obtain more accurate imputation results. Experimental results validate the effectiveness of the proposed method.
Springer Science and Business Media LLC
Title: Self-Supervised Multi-Level Generative Adversarial Network Data Imputation Algorithm
Description:
Abstract Data missing has always been a challenging problem in machine learning.
The Generative Adversarial Imputation Networks (GAIN) have been shown to outperform many existing solutions.
However, in GAIN, because missing values lack ground truth as supervision, it is unable to construct reconstruction loss for missing values and can only judge the reasonableness of imputed values based on reconstruction loss of non-missing values and adversarial loss.
From the perspective of granular computing, data has levels, and data at different levels of granularity encapsulates different knowledge.
Therefore, based on granular computing, this paper proposes a self-supervised multi-level generative adversarial network data imputation algorithm (MGAIN).
Firstly, multiple levels of data are constructed using nested feature set sequences.
Then, GAIN is used to impute missing values at the coarsest granularity level, and the imputation results of missing values at the coarse granularity level are used as supervision for imputing missing values at the fine granularity level, constructing reconstruction loss for missing values at the fine granularity level.
Finally, based on reconstruction loss of missing values, reconstruction loss of non-missing values, and adversarial loss, data at the finer granularity level is imputed.
MGAIN imputes missing values level by level from the coarse granularity level to the fine granularity level to obtain more accurate imputation results.
Experimental results validate the effectiveness of the proposed method.

Related Results

Is a Fitbit a Diary? Self-Tracking and Autobiography
Is a Fitbit a Diary? Self-Tracking and Autobiography
Data becomes something of a mirror in which people see themselves reflected. (Sorapure 270)In a 2014 essay for The New Yorker, the humourist David Sedaris recounts an obsession spu...
Evaluation of sequencing strategies for whole-genome imputation with hybrid peeling
Evaluation of sequencing strategies for whole-genome imputation with hybrid peeling
Abstract Background For assembling large whole-genome sequence datasets to be used routinely in research and breeding, the sequ...
ProDef-MDS: A Proactive Defense Mechanism Protecting Malware Detection Systems from Adversarial Attacks
ProDef-MDS: A Proactive Defense Mechanism Protecting Malware Detection Systems from Adversarial Attacks
Malware threatens cybersecurity by enabling data theft, unauthorized access, and extortion. Traditional malware detection systems (MDS) struggle with the increasing volume and comp...
A Coalescent Model for Genotype Imputation
A Coalescent Model for Genotype Imputation
AbstractThe potential for imputed genotypes to enhance an analysis of genetic data depends largely on the accuracy of imputation, which in turn depends on properties of the referen...
Imputation of Spatially-resolved Transcriptomes by Graph-regularized Tensor Completion
Imputation of Spatially-resolved Transcriptomes by Graph-regularized Tensor Completion
Abstract High-throughput spatial-transcriptomics RNA sequencing (sptRNA-seq) based on in-situ capturing technologies has recently been developed ...
Genotype Imputation
Genotype Imputation
Abstract A missing data problem arises in genetic epidemiological studies when genotypes of particular markers are unavailable fo...
GSimp: A Gibbs sampler based left-censored missing value imputation approach for metabolomics studies
GSimp: A Gibbs sampler based left-censored missing value imputation approach for metabolomics studies
Abstract Left-censored missing values commonly exist in targeted metabolomics datasets and can be considered as missing not at random (MNAR). Imp...
A New Approach of Outlier-robust Missing Value Imputation for Metabolomics Data Analysis
A New Approach of Outlier-robust Missing Value Imputation for Metabolomics Data Analysis
Background:Metabolomics data generation and quantification are different from other types of molecular “omics” data in bioinformatics. Mass spectrometry (MS) based (gas chromatogra...

Back to Top