科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ medRxiv : the preprint server for health sciences2026-07-29

Learning Ophthalmologist Clinical Reasoning for Glaucoma Diagnosis from Fundus Images.

Kaichen Zhou, Yuzhen Chen, Elif Yildiz, Min Shi, David Dai, Grace Chen, Jiale Zheng, He Wang, Fangneng Zhan, Chhavi Saini, Lucy Q Shen, Yike Guo, Paul Pu Liang, Mengyu Wang

原始摘要(英文原文)· Original abstract
Glaucoma is a leading cause of irreversible blindness worldwide. Ophthalmologists diagnose glaucoma through a structured reasoning process by sequentially evaluating optic nerve head characteristics before reaching a final diagnosis, whereas existing AI systems typically perform direct image classification without providing clinically meaningful reasoning. We present the first clinically annotated fundus reasoning dataset, comprising 1,077 fundus photographs paired with expert-authored six-step diagnostic reports. Building on this dataset, we develop a reasoning-driven vision-language framework that explicitly models the ophthalmologist's diagnostic workflow by generating structured clinical reasoning prior to diagnosis. The generated reports are clinically validated, achieving the best performance across all evaluated clinical findings, including a cup-to-disc ratio mean absolute error of 0.070, an ISNT Kendall distance of 1.73, and the highest semantic agreement with expert reports (BERTScore-F1 = 0.874). The resulting framework also improves glaucoma diagnosis, achieving a balanced accuracy of 94.7% and precision of 94.8%, demonstrating that explicitly modeling expert clinical reasoning simultaneously improves interpretability and diagnostic performance. Code and data are available at https://glaucoma-cot.github.io/ .
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Learning Ophthalmologist Clinical Reasoning for Glaucoma Diagnosis from Fundus Images. — 科研速览 Science Skim