Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Automated YOLO-Based Cephalometric Landmark Detection for ANB-Based Skeletal Classification: A Retrospective Single-Centre Study

View through CrossRef
Background/Objectives: Automated cephalometric landmark detection using deep learning has the potential to streamline routine orthodontic diagnosis. However, the clinical relevance of artificial intelligence (AI) localisation accuracy depends on how detection errors propagate into derived angular measurements and skeletal classifications. We retrospectively evaluated 14 YOLO-based model configurations and quantified the agreement between AI-derived and expert-derived ANB-based skeletal classifications. Methods: Twelve working YOLO-based models (YOLOv5xu, YOLOv11 nano/small/medium/large variants) were trained on a single-centre dataset of 120 lateral cephalograms and evaluated on an independent test set of 11 cephalograms (stratified across skeletal Classes I, II, III). The four ANB-defining landmarks (Sella, Nasion, A-point, B-point) were the focus of the analysis. Each test cephalogram had been annotated by four orthodontists (44 measurements per image), yielding the expert reference. We assessed the effects of architecture, bounding-box size (40/100/150 px), training dataset scale (235–4255 images) and training epochs on localisation accuracy (mean radial error, MRE; Successful Detection Rate, SDR) and on the downstream ANB-based skeletal classification. Diagnostic concordance was quantified by classification agreement, Cohen’s κ with bootstrap 95% confidence intervals (10,000 iterations), an exact one-sided binomial test for discordance, and Wilson exact CIs per class. Results: The best-performing model (Model 2; YOLOv11l, 40 × 40 px bounding box, 1175 training images) achieved an MRE of 3.10±1.00 mm and a SDR@4 mm of 87.2% for S, N, A, and B. ANB-based skeletal classification demonstrated 96.9% concordance with expert assessments (95% bootstrap CI: 93.8–99.2%; Cohen’s κ = 0.946 [95% CI 0.89–0.99]; exact binomial test against a 90% concordance threshold p=0.003). Per-class concordance was Class I 95.8% (23/24), Class II 94.9% (56/59), and Class III 100% (47/47). Three of four discordant cases clustered near the Class I/II diagnostic threshold (expert ANB ≈4.5°). Bounding-box size dominated localisation accuracy, with a 3.5-fold increase in MRE from 40 × 40 to 150 × 150 px configurations and SDR@4 mm collapsing from 82.8% to 0%. Conclusions: Within the constraints of a retrospective single-centre design with a small (n = 11) independent test set, YOLO-based AI landmark detection demonstrated promising diagnostic concordance with expert consensus for ANB-based skeletal classification. These findings warrant prospective, multi-centre external validation before clinical deployment and support a confidence-aware workflow in which AI predictions for borderline ANB values undergo mandatory clinician verification. Bounding-box calibration emerged as the single most impactful preprocessing decision.
Title: Automated YOLO-Based Cephalometric Landmark Detection for ANB-Based Skeletal Classification: A Retrospective Single-Centre Study
Description:
Background/Objectives: Automated cephalometric landmark detection using deep learning has the potential to streamline routine orthodontic diagnosis.
However, the clinical relevance of artificial intelligence (AI) localisation accuracy depends on how detection errors propagate into derived angular measurements and skeletal classifications.
We retrospectively evaluated 14 YOLO-based model configurations and quantified the agreement between AI-derived and expert-derived ANB-based skeletal classifications.
Methods: Twelve working YOLO-based models (YOLOv5xu, YOLOv11 nano/small/medium/large variants) were trained on a single-centre dataset of 120 lateral cephalograms and evaluated on an independent test set of 11 cephalograms (stratified across skeletal Classes I, II, III).
The four ANB-defining landmarks (Sella, Nasion, A-point, B-point) were the focus of the analysis.
Each test cephalogram had been annotated by four orthodontists (44 measurements per image), yielding the expert reference.
We assessed the effects of architecture, bounding-box size (40/100/150 px), training dataset scale (235–4255 images) and training epochs on localisation accuracy (mean radial error, MRE; Successful Detection Rate, SDR) and on the downstream ANB-based skeletal classification.
Diagnostic concordance was quantified by classification agreement, Cohen’s κ with bootstrap 95% confidence intervals (10,000 iterations), an exact one-sided binomial test for discordance, and Wilson exact CIs per class.
Results: The best-performing model (Model 2; YOLOv11l, 40 × 40 px bounding box, 1175 training images) achieved an MRE of 3.
10±1.
00 mm and a SDR@4 mm of 87.
2% for S, N, A, and B.
ANB-based skeletal classification demonstrated 96.
9% concordance with expert assessments (95% bootstrap CI: 93.
8–99.
2%; Cohen’s κ = 0.
946 [95% CI 0.
89–0.
99]; exact binomial test against a 90% concordance threshold p=0.
003).
Per-class concordance was Class I 95.
8% (23/24), Class II 94.
9% (56/59), and Class III 100% (47/47).
Three of four discordant cases clustered near the Class I/II diagnostic threshold (expert ANB ≈4.
5°).
Bounding-box size dominated localisation accuracy, with a 3.
5-fold increase in MRE from 40 × 40 to 150 × 150 px configurations and SDR@4 mm collapsing from 82.
8% to 0%.
Conclusions: Within the constraints of a retrospective single-centre design with a small (n = 11) independent test set, YOLO-based AI landmark detection demonstrated promising diagnostic concordance with expert consensus for ANB-based skeletal classification.
These findings warrant prospective, multi-centre external validation before clinical deployment and support a confidence-aware workflow in which AI predictions for borderline ANB values undergo mandatory clinician verification.
Bounding-box calibration emerged as the single most impactful preprocessing decision.

Related Results

YOLO-Based Automated Cephalometric Landmark Detection Achieves Expert-Level Skeletal Classification: A Clinical Validation Study
YOLO-Based Automated Cephalometric Landmark Detection Achieves Expert-Level Skeletal Classification: A Clinical Validation Study
Automated cephalometric landmark detection using deep learning has the potential to transform routine orthodontic diagnosis. However, the clinical relevance of AI localization accu...
Lightweight fruit detection algorithms for low‐power computing devices
Lightweight fruit detection algorithms for low‐power computing devices
Abstract A lightweight fruit detection algorithm is important to ensure real‐time detection on low‐power computing devices while maintaining detection accuracy. I...
Are Cervical Ribs Indicators of Childhood Cancer? A Narrative Review
Are Cervical Ribs Indicators of Childhood Cancer? A Narrative Review
Abstract A cervical rib (CR), also known as a supernumerary or extra rib, is an additional rib that forms above the first rib, resulting from the overgrowth of the transverse proce...
Application of YOLO-v7 and YOLO-v8 Transfer Learning Models in Breast Lesion Classification and Diagnosis
Application of YOLO-v7 and YOLO-v8 Transfer Learning Models in Breast Lesion Classification and Diagnosis
Background: Early detection of breast cancer and accurate assessment of lesions are key goals of imaging evaluation. Ultrasound is widely used, but its diagnost...
P0553PROTECTIVE EFFECT OF (-)-ALPHA-BISABOLOL OVER RENAL CYTOTOXICITY INDUCED BY AMPHOTERICIN B
P0553PROTECTIVE EFFECT OF (-)-ALPHA-BISABOLOL OVER RENAL CYTOTOXICITY INDUCED BY AMPHOTERICIN B
Abstract Background and Aims Drug-induced acute kidney injury (AKI) is a health problem that increases morbidity and mortality r...
Osoby niejednokrotnie przebywające w izbie wytrzeźwień
Osoby niejednokrotnie przebywające w izbie wytrzeźwień
In Poland we have at present in towns 29 detoxication centres with 1,226 beds; people found by the police in public places in a state of intoxication are more and more often taken ...

Back to Top