Post

New research · Ophthalmology
Frontiers in endocrinology · 4d
AI / informaticsFrontiers in endocrinology · 2026

Diagnostic accuracy and clinical performance of deep learning models for grading diabetic retinopathy: a systematic review and meta-analysis.

Xin Yan, Shiqi Lei, Lifen Hu … Na Wu
Read paper
OphthalmologyAI / informatics

Deep learning detects no diabetic retinopathy with very high sensitivity on fundus photos

Diagnostic accuracy and clinical performance of deep learning models for grading diabetic retinopathy: a systematic review and meta-analysis.

Xin Yan … Na Wu
Frontiers in endocrinology · 2026
Background

Diabetic retinopathy (diabetic retinopathy) is a leading cause of preventable visual impairment worldwide, and its precise severity grading is critical for optimizing clinical management.

Purpose

This systematic review and meta-analysis aimed to comprehensively assess the diagnostic accuracy of fundus image-based deep learning models in the grading of diabetic retinopathy.

Methods

PubMed, Embase, Web of Science, and the Cochrane Library were systematically searched for relevant studies published up to October 28, 2025.

n = 41 studies
Results

correctly flags nearly all eyes that truly have no diabetic retinopathy

No DR (stage 0)
95.19%
Mild NPDR (stage 1)
72.06%
Moderate NPDR (stage 2
84.33%
Severe NPDR (stage 3)
75.84%
More results

In the simplified four-class classification task, sensitivities markedly improved across all grades: 96.85% (95% CI: 90.18%-99.93%) for stage 0, 92.94% (95% CI: 79.50%-99.72%) for stage 1, 92.75% (95% CI: 79.31%-99.61%) for stage 2, and 88.19% (95% CI: 68.99%-98.93%) for stage 3.

“
Conclusion · 1 of 3

Deep learning exhibits high sensitivity and substantial potential for diabetic retinopathy grading, particularly in screening for no diabetic retinopathy and vision-threatening diabetic retinopathy.

Conclusion · 2 of 3

Nevertheless, precisely differentiating between adjacent non-proliferative stages remains a clinical challenge.

Conclusion · 3 of 3

The observed heterogeneity underscores the imperative for methodological standardization, rigorous external validation, and multimodal data integration.

Read paper
0 comments

No comments yet. Be the first.

Related papers

LatestFoundational
AI / Informatics
0·962 for OCT
MerMED-FM was very accurate at diagnosing diseases using eye scans
AI / Informatics
0.98
Artificial intelligence's eyelid-height measurements matched doctors' manual measurements almost perfectly
AI / Informatics
92.1%
combining eye scans and photos correctly told benign from cancerous lesions apart nearly every time
Cohort Study
28.6%
recurred locally in more than 1 in 4 patients over years of follow-up
Cohort Study
-0.395
excision group's post-op eyelid fullness score was lower - greater improvement
Cohort Study
86.7%
of infants probed after 12 months still had unresolved tear duct blockage
Observational
16%
eyelid tissue in rosacea patients showed less of this key repair-signaling protein inside cell nuclei
Cohort Study
25%
about 1 in 4 treated cases had symptoms return after improving