Javascript must be enabled to continue!
Label Ranker: Self-Aware Preference for Classification Label Position in Visual Masked Self-Supervised Pre-Trained Model
View through CrossRef
This paper investigates the impact of randomly initialized unique encoding of classification label position on the visual masked self-supervised pre-trained model when fine-tuning downstream classification tasks. Our findings indicate that different random initializations lead to significant variations in fine-tuned results, even when using the same allocation strategy for classification datasets. The accuracy gap between these results suggests that the visual masked self-supervised pre-trained model has an inherent preference for classification label positions. To investigate this, we compare it with the non-self-supervised visual pre-trained model and hypothesize that the masked self-supervised model exhibits a self-aware bias toward certain label positions. To mitigate the instability caused by random encoding, we propose a classification label position ranking algorithm, Label Ranker. It is based on 1-D dimensionality reduction of feature maps using Linear Discriminant Analysis and position-rank encoding of them by unsupervised feature clustering using the similarity property of Euclidean distance. This algorithm ensures that label position encoding align with the model is inherent preference. Extensive ablation experiments using ImageMAE and VideoMAE models on the CIFAR-100, UCF101, and HMDB51 classification datasets validate our approach. Results demonstrate that our method effectively stabilizes classification label position encoding, improving fine-tuned performance for visual masked self-supervised models.
Title: Label Ranker: Self-Aware Preference for Classification Label Position in Visual Masked Self-Supervised Pre-Trained Model
Description:
This paper investigates the impact of randomly initialized unique encoding of classification label position on the visual masked self-supervised pre-trained model when fine-tuning downstream classification tasks.
Our findings indicate that different random initializations lead to significant variations in fine-tuned results, even when using the same allocation strategy for classification datasets.
The accuracy gap between these results suggests that the visual masked self-supervised pre-trained model has an inherent preference for classification label positions.
To investigate this, we compare it with the non-self-supervised visual pre-trained model and hypothesize that the masked self-supervised model exhibits a self-aware bias toward certain label positions.
To mitigate the instability caused by random encoding, we propose a classification label position ranking algorithm, Label Ranker.
It is based on 1-D dimensionality reduction of feature maps using Linear Discriminant Analysis and position-rank encoding of them by unsupervised feature clustering using the similarity property of Euclidean distance.
This algorithm ensures that label position encoding align with the model is inherent preference.
Extensive ablation experiments using ImageMAE and VideoMAE models on the CIFAR-100, UCF101, and HMDB51 classification datasets validate our approach.
Results demonstrate that our method effectively stabilizes classification label position encoding, improving fine-tuned performance for visual masked self-supervised models.
Related Results
Self-Supervised Transformer Networks: Unlocking New Possibilities for Label-Free Data
Self-Supervised Transformer Networks: Unlocking New Possibilities for Label-Free Data
In machine learning, self-supervised transformer networks have become a new way of doing things, especially when it comes to handling and understanding huge amounts of data that ha...
Hypertension-mediated organ damage in masked hypertension
Hypertension-mediated organ damage in masked hypertension
Objectives:
Masked hypertension – a blood pressure (BP) phenotype characterized by a clinic BP in the normal range but elevated BP outside the office – is associated wi...
Automatic classification of construction accident reports using BERTopic-GLDA approach
Automatic classification of construction accident reports using BERTopic-GLDA approach
Purpose
This study aims to propose a semi-supervised classification framework that reduces reliance on labeled data, manages class imbalance and improves the in...
Is a Fitbit a Diary? Self-Tracking and Autobiography
Is a Fitbit a Diary? Self-Tracking and Autobiography
Data becomes something of a mirror in which people see themselves reflected. (Sorapure 270)In a 2014 essay for The New Yorker, the humourist David Sedaris recounts an obsession spu...
Generalization of vision pre-trained models for histopathology
Generalization of vision pre-trained models for histopathology
AbstractOut-of-distribution (OOD) generalization, especially for medical setups, is a key challenge in modern machine learning which has only recently received much attention. We i...
Self-supervised pre-training with contrastive and masked autoencoder methods for dealing with small datasets in deep learning for medical imaging
Self-supervised pre-training with contrastive and masked autoencoder methods for dealing with small datasets in deep learning for medical imaging
AbstractDeep learning in medical imaging has the potential to minimize the risk of diagnostic errors, reduce radiologist workload, and accelerate diagnosis. Training such deep lear...
Pre-trained Vision Transformer With Masked Autoencoder for Automated Diabetic Macular Edema Detection from Optical Coherence Tomography Images
Pre-trained Vision Transformer With Masked Autoencoder for Automated Diabetic Macular Edema Detection from Optical Coherence Tomography Images
Abstract
Purpose
To develop and evaluate a novel self-supervised learning approach using Masked Autoencoder (MAE) pre-trained V...
Cardiovascular Risk Factors and Masked Hypertension
Cardiovascular Risk Factors and Masked Hypertension
Masked hypertension is associated with increased risk for cardiovascular disease. Identifying modifiable risk factors for masked hypertension could provide approaches to reduce its...

