Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Systematic Review of Hybrid Vision Transformer Architectures for Radiological Image Analysis

View through CrossRef
Background Vision Transformer (ViT) and Convolutional Neural Networks (CNNs) each possess distinct strengths in medical imaging: ViT excels in capturing long-range dependencies through self-attention, while CNNs are adept at extracting local features via spatial convolution filters. However, ViT may struggle with detailed local spatial information, critical for tasks like anomaly detection in medical imaging, while shallow CNNs may not effectively abstract global context. Objective This study aims to explore and evaluate hybrid architectures that integrate ViT and CNN to lever-age their complementary strengths for enhanced performance in medical vision tasks, such as segmentation, classification, and prediction. Methods Following PRISMA guidelines, a systematic review was conducted on 28 articles published between 2020 and 2023. These articles proposed hybrid ViT-CNN architectures specifically for medical imaging tasks in radiology. The review focused on analyzing architectural variations, merging strategies between ViT and CNN, innovative applications of ViT, and efficiency metrics including parameters, inference time (GFlops), and performance benchmarks. Results The review identified that integrating ViT and CNN can mitigate the limitations of each architecture, offering comprehensive solutions that combine global context understanding with precise local feature extraction. We benchmarked the articles based on architectural variations, merging strategies, innovative uses of ViT, and efficiency metrics (number of parameters, inference time(GFlops), performance). Conclusion By synthesizing current literature, this review defines fundamental concepts of hybrid vision transformers and highlights emerging trends in the field. It provides a clear direction for future research aimed at optimizing the integration of ViT and CNN for effective utilization in medical imaging, contributing to advancements in diagnostic accuracy and image analysis. Summary Statement We performed systematic review of hybrid vision transformer architecture using PRISMA guideline and perfromed through meta-analysis to benchmark the architectures. ACM Reference Format Ji Woong Kim, Aisha Urooj Khan, and Imon Banerjee. 2018. Systematic Review of Hybrid Vision Transformer Architectures for Radiological Image Analysis. J. ACM 37, 4, Article 111 (August 2018), 16 pages. https://doi.org/XXXXXXX.XXXXXXX
Title: Systematic Review of Hybrid Vision Transformer Architectures for Radiological Image Analysis
Description:
Background Vision Transformer (ViT) and Convolutional Neural Networks (CNNs) each possess distinct strengths in medical imaging: ViT excels in capturing long-range dependencies through self-attention, while CNNs are adept at extracting local features via spatial convolution filters.
However, ViT may struggle with detailed local spatial information, critical for tasks like anomaly detection in medical imaging, while shallow CNNs may not effectively abstract global context.
Objective This study aims to explore and evaluate hybrid architectures that integrate ViT and CNN to lever-age their complementary strengths for enhanced performance in medical vision tasks, such as segmentation, classification, and prediction.
Methods Following PRISMA guidelines, a systematic review was conducted on 28 articles published between 2020 and 2023.
These articles proposed hybrid ViT-CNN architectures specifically for medical imaging tasks in radiology.
The review focused on analyzing architectural variations, merging strategies between ViT and CNN, innovative applications of ViT, and efficiency metrics including parameters, inference time (GFlops), and performance benchmarks.
Results The review identified that integrating ViT and CNN can mitigate the limitations of each architecture, offering comprehensive solutions that combine global context understanding with precise local feature extraction.
We benchmarked the articles based on architectural variations, merging strategies, innovative uses of ViT, and efficiency metrics (number of parameters, inference time(GFlops), performance).
Conclusion By synthesizing current literature, this review defines fundamental concepts of hybrid vision transformers and highlights emerging trends in the field.
It provides a clear direction for future research aimed at optimizing the integration of ViT and CNN for effective utilization in medical imaging, contributing to advancements in diagnostic accuracy and image analysis.
Summary Statement We performed systematic review of hybrid vision transformer architecture using PRISMA guideline and perfromed through meta-analysis to benchmark the architectures.
ACM Reference Format Ji Woong Kim, Aisha Urooj Khan, and Imon Banerjee.
2018.
Systematic Review of Hybrid Vision Transformer Architectures for Radiological Image Analysis.
J.
ACM 37, 4, Article 111 (August 2018), 16 pages.
https://doi.
org/XXXXXXX.
XXXXXXX.

Related Results

Evaluating the Science to Inform the Physical Activity Guidelines for Americans Midcourse Report
Evaluating the Science to Inform the Physical Activity Guidelines for Americans Midcourse Report
Abstract The Physical Activity Guidelines for Americans (Guidelines) advises older adults to be as active as possible. Yet, despite the well documented benefits of physical activi...
Automatic Load Sharing of Transformer
Automatic Load Sharing of Transformer
Transformer plays a major role in the power system. It works 24 hours a day and provides power to the load. The transformer is excessive full, its windings are overheated which lea...
Depth-aware salient object segmentation
Depth-aware salient object segmentation
Object segmentation is an important task which is widely employed in many computer vision applications such as object detection, tracking, recognition, and ret...
High frequency modeling of power transformers under transients
High frequency modeling of power transformers under transients
This thesis presents the results related to high frequency modeling of power transformers. First, a 25kVA distribution transformer under lightning surges is tested in the laborator...
Do evidence summaries increase health policy‐makers' use of evidence from systematic reviews? A systematic review
Do evidence summaries increase health policy‐makers' use of evidence from systematic reviews? A systematic review
This review summarizes the evidence from six randomized controlled trials that judged the effectiveness of systematic review summaries on policymakers' decision making, or the most...
Lists, Spatial Practice and Assistive Technologies for the Blind
Lists, Spatial Practice and Assistive Technologies for the Blind
IntroductionSupermarkets are functionally challenging environments for people with vision impairments. A supermarket is likely to house an average of 45,000 products in a median fl...
Hydatid Disease of The Brain Parenchyma: A Systematic Review
Hydatid Disease of The Brain Parenchyma: A Systematic Review
Abstarct Introduction Isolated brain hydatid disease (BHD) is an extremely rare form of echinococcosis. A prompt and timely diagnosis is a crucial step in disease management. This ...
ANALISIS PENGARUH MASA OPERASIONAL TERHADAP PENURUNAN KAPASITAS TRANSFORMATOR DISTRIBUSI DI PT PLN (PERSERO)
ANALISIS PENGARUH MASA OPERASIONAL TERHADAP PENURUNAN KAPASITAS TRANSFORMATOR DISTRIBUSI DI PT PLN (PERSERO)
One cause the interruption of transformer is loading that exceeds the capabilities of the transformer. The state of continuous overload will affect the age of the transformer and r...

Back to Top