Javascript must be enabled to continue!
Pre-trained Vision Transformer With Masked Autoencoder for Automated Diabetic Macular Edema Detection from Optical Coherence Tomography Images
View through CrossRef
Abstract
Purpose
To develop and evaluate a novel self-supervised learning approach using Masked Autoencoder (MAE) pre-trained Vision Transformer (ViT) for automated detection of diabetic macular edema (DME) from optical coherence tomography (OCT) images, addressing the critical need for scalable screening solutions in diabetic eye care.
Study Design
Artificial intelligence model training.
Methods
We utilized the publicly available Kermany dataset containing 109,312 OCT images, defining DME detection as a binary classification task (11,559 DME vs. 97,753 non-DME images). Five deep learning architectures were compared: MAE-pretrained ViT (MAE_ViT), standard ViT, ResNet18, VGG19_bn, and EfficientNetV2. MAE_ViT underwent two-stage training: (1) self-supervised pre-training with 75% patch masking for 1,000 epochs to learn robust visual representations, and (2) supervised fine-tuning for DME classification. Model performance was evaluated using accuracy, sensitivity, specificity, F1 score, and area under the receiver operating characteristic curve (AU-ROC) with 95% confidence intervals calculated via bootstrap resampling.
Results
MAE_ViT achieved superior performance with AU-ROC 0.999 (95% CI: 0.999-1.000), accuracy 98.5% (95% CI: 97.7-99.2%), sensitivity 99.6% (95% CI: 98.7-100%), and specificity 98.1% (95% CI: 97.2-99.1%). VGG19_bn showed the second-best performance (AU-ROC 0.997), while ResNet18 demonstrated poor specificity (28.3%) despite perfect sensitivity. The self-supervised approach of MAE_ViT outperformed standard supervised ViT (AU-ROC 0.995), demonstrating the effectiveness of learning from unlabeled data.
Conclusion
MAE pre-trained Vision Transformer establishes a new benchmark for automated DME detection, offering exceptional diagnostic accuracy and potential for deployment in resource-constrained settings through reduced annotation requirements.
Title: Pre-trained Vision Transformer With Masked Autoencoder for Automated Diabetic Macular Edema Detection from Optical Coherence Tomography Images
Description:
Abstract
Purpose
To develop and evaluate a novel self-supervised learning approach using Masked Autoencoder (MAE) pre-trained Vision Transformer (ViT) for automated detection of diabetic macular edema (DME) from optical coherence tomography (OCT) images, addressing the critical need for scalable screening solutions in diabetic eye care.
Study Design
Artificial intelligence model training.
Methods
We utilized the publicly available Kermany dataset containing 109,312 OCT images, defining DME detection as a binary classification task (11,559 DME vs.
97,753 non-DME images).
Five deep learning architectures were compared: MAE-pretrained ViT (MAE_ViT), standard ViT, ResNet18, VGG19_bn, and EfficientNetV2.
MAE_ViT underwent two-stage training: (1) self-supervised pre-training with 75% patch masking for 1,000 epochs to learn robust visual representations, and (2) supervised fine-tuning for DME classification.
Model performance was evaluated using accuracy, sensitivity, specificity, F1 score, and area under the receiver operating characteristic curve (AU-ROC) with 95% confidence intervals calculated via bootstrap resampling.
Results
MAE_ViT achieved superior performance with AU-ROC 0.
999 (95% CI: 0.
999-1.
000), accuracy 98.
5% (95% CI: 97.
7-99.
2%), sensitivity 99.
6% (95% CI: 98.
7-100%), and specificity 98.
1% (95% CI: 97.
2-99.
1%).
VGG19_bn showed the second-best performance (AU-ROC 0.
997), while ResNet18 demonstrated poor specificity (28.
3%) despite perfect sensitivity.
The self-supervised approach of MAE_ViT outperformed standard supervised ViT (AU-ROC 0.
995), demonstrating the effectiveness of learning from unlabeled data.
Conclusion
MAE pre-trained Vision Transformer establishes a new benchmark for automated DME detection, offering exceptional diagnostic accuracy and potential for deployment in resource-constrained settings through reduced annotation requirements.
Related Results
ASSOCIATION BETWEEN THE DEGREE OF DIABETIC RETINOPATHY AND DIABETIC MACULAR EDEMA
ASSOCIATION BETWEEN THE DEGREE OF DIABETIC RETINOPATHY AND DIABETIC MACULAR EDEMA
INTRODUCTION: WHO estimates more than 150 million diabetes patients worldwide. One of the complications of diabetes is diabetic retinopathy which is recognized as the leading cause...
Association of Types of Diabetic Macular Edema with Different Anti-Diabetic Therapies
Association of Types of Diabetic Macular Edema with Different Anti-Diabetic Therapies
ABSTRACT
Objective: To evaluate and assess the association of diabetic macular edema with different anti-diabetic therapy regimens. Material and Methods: We recruited 340 patients...
EFFECT OF PHACOEMULSIFICATION ON THE PROGRESSION OF DIABETIC RETINOPATHY AND DIABETIC MACULAR EDEMA
EFFECT OF PHACOEMULSIFICATION ON THE PROGRESSION OF DIABETIC RETINOPATHY AND DIABETIC MACULAR EDEMA
Background: DR is the leading cause of vision loss in adults aged 22-74 years. DR ranked as the fth most common causes of preventable
blindness. In cataract surgery, blood aqueous...
Study of Role of Optical Coherence Tomography in the Interpretation of Common Macular Disorders at a Tertiary Eye Care Centre in Rural Maharashtra
Study of Role of Optical Coherence Tomography in the Interpretation of Common Macular Disorders at a Tertiary Eye Care Centre in Rural Maharashtra
BACKGROUND Optical Coherence Tomography (OCT) is a commonly used non invasive imaging instrument useful for the diagnosis and follow up of macular disorders, but it has its own sha...
Latent Diabetic Macular Edema in Chinese Diabetic Retinopathy Patients
Latent Diabetic Macular Edema in Chinese Diabetic Retinopathy Patients
Purpose: To compare the detection rates of optical coherence tomography (OCT) and fluorescein angiography (FA) in a diabetic macular edema (DME) and the severity of diabetic retino...
Prognostic Role of Optical Coherence Tomography in Outcome of Idiopathic full thickness Macular Hole Surgery
Prognostic Role of Optical Coherence Tomography in Outcome of Idiopathic full thickness Macular Hole Surgery
Abstract
Purpose: To assess the prognostic value of Optical Coherence Tomography indices preoperatively in outcome of idiopathic full thickness macular hole surgery.
Ma...
Automatic Load Sharing of Transformer
Automatic Load Sharing of Transformer
Transformer plays a major role in the power system. It works 24 hours a day and provides power to the load. The transformer is excessive full, its windings are overheated which lea...
Macular Thickness Decreases with Age in Normal Eyes: A Study On Macular Thickness Map Protocol on the Oct
Macular Thickness Decreases with Age in Normal Eyes: A Study On Macular Thickness Map Protocol on the Oct
Objective: To evaluate the normal macular thickness and how it varies with age in the eyes of healthy subjects using macular map protocol on Spectral-Domain Optical Coherence Tomog...

