[关键词]
[摘要]
随着大模型的发展,通用和眼科专用大模型在眼科图像分析中展现出巨大潜力,但其在特定临床任务中的应用效能及适用边界尚待明确。文章梳理了眼科大模型的技术路线演进,包括掩码视觉建模、图文对比学习以及知识融合与多模态推理等阶段,并对比分析了通用模型与专用模型在疾病筛查与初诊、疾病分级与病灶分割、医疗报告生成与多模态推理等常见临床场景中的性能差异。结果显示,通用模型在多病种初筛中表现较好,而眼科专用模型在精细化分级、微小病灶识别及罕见病诊断等任务中占据优势,且经指令微调后生成的医疗报告更具临床一致性,这些结果为具体临床任务中的模型类型及技术路线选择提供了参考。此外,文章进一步指出了当前模型面临的可靠性、泛化性、评估体系及可重复性挑战及其所造成的临床应用局限性。
[Key word]
[Abstract]
With the rapid advancement of large language models, both general-purpose and ophthalmology-specific variants have shown substantial potential for ophthalmic image analysis. However, their task-specific effectiveness and applicable boundaries in clinical practice remain to be fully clarified. This review outlines the technical evolution of large models in ophthalmology, including masked visual modeling, vision-language contrastive learning, and subsequent stages of knowledge integration and multimodal reasoning. The performance of general-purpose versus ophthalmology-specific models across common clinical scenarios, including disease screening and initial triage, disease grading and lesion segmentation, and medical report generation with multimodal reasoning were further compared. Results show that general-purpose models are well-suited for multi-disease screening and primary-care triage, whereas ophthalmology-specific models excel in fine-grained grading, subtle lesions detection, and rare disease diagnosis. Moreover, instruction-tuned domain-specific models produce medical reports with greater clinical consistency. These findings provide practical guidance for selecting appropriate model types and technical pathways tailored to specific ophthalmic tasks. Furthermore, the review identifies the current challenges facing these models, including reliability, generalizability, evaluation frameworks, and reproducibility, as well as the resulting limitations in clinical application.
[中图分类号]
[基金项目]
广东省自然科学基金面上项目(No.2024A1515012292); 中山大学中山眼科中心优秀科技创新人才支持与培育计划项目(No.CXQN-010)