Eye and ENT hospital of Fudan University
Shanghai, Shanghai Municipality, 200000, China
Location status: Recruiting
NCT Number: NCT07183891
We conducted a single-center, retrospective observational study to evaluate large language models (ChatGPT 4o, GPT-5, DeepSeek) for automated interpretation of de-identified IOLMaster 700 reports provided as raster images. Models produced structured biometric extraction, toric IOL recommendation, and refractive predictions (sphere, cylinder, axis). Primary outcomes included parameter-level agreement and refractive error metrics; secondary outcomes included decision-support performance for toric IOL selection and agreement on ordered T-codes. No clinical intervention was performed.
Interested in participating?
Request Info18 year and older
All sexes
Observational
Shanghai, Shanghai Municipality, 200000, China
Location status: Recruiting
This study compares three large language models accessed in their native configurations, without fine-tuning or external tools. For each examination, the original IOLMaster 700 report image was supplied without manual annotation or pre-processing. A standardized instruction required: (i) structured extraction of AL, ACD, LT, WTW, K1/K2 and axes, ΔK, TK1/TK2 and axes, and ΔTK; (ii) binary toric candidacy and T-code according to institutional ALCON mapping; and (iii) refractive recommendations (sphere, cylinder, implantation axis). Each model generated three independent outputs per case. De-identification and IRB oversight (waiver of consent) were implemented according to institutional policy. The unit of enrollment is participants (n=54), with outcomes analyzed per eye (162 eyes) and per model generation where applicable.
Healthy volunteers accepted: No
Only the study team can determine whether someone qualifies for participation.
Inclusion criteria
-postoperative corrected distance visual acuity (CDVA) of 0.10 logMAR or better -an absolute IOL rotational stability of less than 10∘ at the 1-month follow-up examination
Exclusion criteria
Time frame: At index examination
Mean absolute error (MAE, diopters) of model-predicted sphere versus clinical reference
Time frame: At index examination (single time point)
Cohen's kappa with 95% CIs between model outputs and clinician-validated reference for per-parameter
Time frame: At index examination
Mean absolute error (MAE, diopters) of model-predicted Cylinder
Time frame: At index examination
Mean absolute error (MAE, diopters) of model-predicted Axis
Contact information is provided by the study sponsor or research team.
Jin Yang
Other
Head-to-Head Evaluation of ChatGPT 4o, GPT-5, and DeepSeek for Structured Extraction, Toric IOL Recommendation, and Refractive Prediction
OpenTrials presents study information sourced from ClinicalTrials.gov. The official registry record should be consulted for the latest information.
View the official ClinicalTrials.gov record (opens in a new tab)This listing is for discovery and informational purposes only. It is not medical advice, does not guarantee that a study is recruiting, and does not determine eligibility. Contact the study team and a qualified healthcare professional when considering participation.
Published trials that share one or more normalized conditions with this study.
NCT07657858
Astigmatism, Cataract
Shanghai, Shanghai Municipality, China
View Trial DetailsNCT07453992
Astigmatism, Cataract
Shanghai, Shanghai Municipality, China
View Trial DetailsNCT07232615
Astigmatism, Cataract
Düsseldorf, North Rhine-Westphalia, Germany
View Trial DetailsNCT07081919
Astigmatism, Cataract
Shanghai, Shanghai Municipality, China
View Trial Details