Post

New research · Ophthalmology
Ophthalmology science · 22h
AI / informaticsOphthalmology science · 2026

Large Language Models for Ophthalmology Training in China: A Prospective Evaluation.

Zuhui Zhang, Changke Huang, Xinxin Yu … Qi Dai
Read paper
OphthalmologyAI / informatics

Large language model assistance improved resident physician accuracy on text-based ophthalmology exams.

Large Language Models for Ophthalmology Training in China: A Prospective Evaluation.

Zuhui Zhang et al. · Ophthalmology science · 2026
Purpose

This study explored large language models (LLMs) as a scalable solution to the global shortage and uneven distribution of ophthalmologists, particularly their actual effectiveness and potential risks in ophthalmic training.

Methods

Phase 1: all LLMs were tested on the Chinese and English versions of the Chinese National Health Professional Technical Qualification Examination (Intermediate Level) in Ophthalmology (CNHPTQE-O).

Results

resident physicians' accuracy on text-based eye exams improved with AI assistance

60.75%
Before
79%
After
More results

Several Chinese LLMs, especially ERNIE Bot 4.5 Turbo, demonstrated superior performance on the CNHPTQE-O, achieving accuracies of 98.00% (Chinese) and 86.50% (English).

ERNIE Bot 4.5 Turbo significantly outperformed all RPs on the Chinese examination ( P = 0.001).

Questionnaire feedback was positive.

“
Conclusion · 1 of 2

Large language models possess a solid foundation in ophthalmic knowledge and can effectively enhance trainee performance in text-based assessments, demonstrating clear potential as a training aid.

Conclusion · 2 of 2

However, their limitations in image-assisted diagnostic tasks and the associated risk of "artificial ignorance" should not be overlooked.

Read paper
0 comments

No comments yet. Be the first.

Related papers

LatestFoundational
Cohort Study
median 9
optometry referrals contained more complete documentation for glaucoma
Study
71.27%
medical students' higher accuracy in identifying eye infections after artificial intelligence training
Cross-sectional
64%
most patients reported little difficulty with their eye injections
Cross-sectional
37%
Only 37% of primary care providers and endocrinologists correctly identified diabetic retinopathy status.
Randomized Trial
adjusted difference -4.97
Baduanjin exercise led to a better overall quality of life than routine care
Study
58.26%
Qwen-7B correctly found eye diseases in over half of cases without specific training
Case Report
24 months
the patient's eye cancer remained gone for this period after treatment
Study
Mean 184 vs. 137
Residents in subsidized programs performed more cataract surgeries as primary surgeon.