• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

相似文献

1
Vision-Language Models for Feature Detection of Macular Diseases on Optical Coherence Tomography.基于视觉-语言模型的光学相干断层扫描图像黄斑病变特征检测
JAMA Ophthalmol. 2024 Jun 1;142(6):573-576. doi: 10.1001/jamaophthalmol.2024.1165.
2
Radial versus raster spectral-domain optical coherence tomography scan patterns for detection of macular pathology.用于检测黄斑病变的径向与光栅光谱域光学相干断层扫描模式
Am J Ophthalmol. 2014 Aug;158(2):345-353.e2. doi: 10.1016/j.ajo.2014.05.013. Epub 2014 May 20.
3
High-definition and 3-dimensional imaging of macular pathologies with high-speed ultrahigh-resolution optical coherence tomography.利用高速超高分辨率光学相干断层扫描技术对黄斑病变进行高清三维成像。
Ophthalmology. 2006 Nov;113(11):2054.e1-14. doi: 10.1016/j.ophtha.2006.05.046.
4
[Detection and evaluation of OCT findings in macular diseases].[黄斑疾病中光学相干断层扫描(OCT)结果的检测与评估]
Klin Monbl Augenheilkd. 2010 Aug;227(8):R107-25; quiz R126-7. doi: 10.1055/s-0030-1250269. Epub 2010 Aug 12.
5
Computerized macular pathology diagnosis in spectral domain optical coherence tomography scans based on multiscale texture and shape features.基于多尺度纹理和形状特征的谱域光学相干断层扫描计算机黄斑病变诊断。
Invest Ophthalmol Vis Sci. 2011 Oct 21;52(11):8316-22. doi: 10.1167/iovs.10-7012.
6
Diagnostic Utility of Swept-Source OCT-Based Biometry and Fundus Photographs Compared to Spectral Domain OCT in Center-Involving Diabetic Macular Edema.基于扫频源光学相干断层扫描的生物测量和眼底照片与光谱域光学相干断层扫描在累及黄斑中心的糖尿病性黄斑水肿中的诊断效用比较
Ophthalmic Epidemiol. 2025 Feb;32(1):95-102. doi: 10.1080/09286586.2024.2338824. Epub 2024 May 6.
7
Spectral-domain optical coherence tomography use in macular diseases: a review.谱域光学相干断层扫描在黄斑疾病中的应用:综述。
Ophthalmologica. 2010;224(6):333-40. doi: 10.1159/000313814. Epub 2010 May 4.
8
Fully Automated Detection and Quantification of Macular Fluid in OCT Using Deep Learning.基于深度学习的 OCT 中黄斑区液全自动化检测和定量分析
Ophthalmology. 2018 Apr;125(4):549-558. doi: 10.1016/j.ophtha.2017.10.031. Epub 2017 Dec 8.
9
Optical coherence tomography (OCT) for detection of macular oedema in patients with diabetic retinopathy.光学相干断层扫描(OCT)用于检测糖尿病视网膜病变患者的黄斑水肿。
Cochrane Database Syst Rev. 2015 Jan 7;1(1):CD008081. doi: 10.1002/14651858.CD008081.pub3.
10
Evaluation of an Artificial Intelligence-Based Detector of Sub- and Intraretinal Fluid on a Large Set of Optical Coherence Tomography Volumes in Age-Related Macular Degeneration and Diabetic Macular Edema.基于人工智能的视网膜下和视网膜内液检测在大量年龄相关性黄斑变性和糖尿病性黄斑水肿光学相干断层扫描体积上的评估。
Ophthalmologica. 2022;245(6):516-527. doi: 10.1159/000527345. Epub 2022 Oct 10.

引用本文的文献

1
Large language models for disease diagnosis: a scoping review.用于疾病诊断的大语言模型:一项范围综述。
NPJ Artif Intell. 2025;1(1):9. doi: 10.1038/s44387-025-00011-z. Epub 2025 Jun 9.
2
Can off-the-shelf visual large language models detect and diagnose ocular diseases from retinal photographs?现成的视觉大语言模型能否从视网膜照片中检测和诊断眼部疾病?
BMJ Open Ophthalmol. 2025 Apr 7;10(1):e002076. doi: 10.1136/bmjophth-2024-002076.
3
Large Language Models in Ophthalmology: A Review of Publications from Top Ophthalmology Journals.眼科领域的大语言模型:顶级眼科期刊出版物综述
Ophthalmol Sci. 2024 Dec 17;5(3):100681. doi: 10.1016/j.xops.2024.100681. eCollection 2025 May-Jun.
4
Multimodal machine learning enables AI chatbot to diagnose ophthalmic diseases and provide high-quality medical responses.多模态机器学习使人工智能聊天机器人能够诊断眼科疾病并提供高质量的医学回复。
NPJ Digit Med. 2025 Jan 27;8(1):64. doi: 10.1038/s41746-025-01461-0.
5
Artificial Intelligence for Optical Coherence Tomography in Glaucoma.用于青光眼光学相干断层扫描的人工智能
Transl Vis Sci Technol. 2025 Jan 2;14(1):27. doi: 10.1167/tvst.14.1.27.
6
Evolutionary patterns and research frontiers of artificial intelligence in age-related macular degeneration: a bibliometric analysis.年龄相关性黄斑变性中人工智能的进化模式与研究前沿:一项文献计量分析
Quant Imaging Med Surg. 2025 Jan 2;15(1):813-830. doi: 10.21037/qims-24-1406. Epub 2024 Dec 30.
7
Foundation models in ophthalmology: opportunities and challenges.眼科领域的基础模型:机遇与挑战。
Curr Opin Ophthalmol. 2025 Jan 1;36(1):90-98. doi: 10.1097/ICU.0000000000001091. Epub 2024 Nov 4.
8
Capabilities of GPT-4o and Gemini 1.5 Pro in Gram stain and bacterial shape identification.GPT-4o 和 Gemini 1.5 Pro 在革兰氏染色和细菌形态识别方面的能力。
Future Microbiol. 2024;19(15):1283-1292. doi: 10.1080/17460913.2024.2381967. Epub 2024 Jul 29.
9
Generative artificial intelligence in ophthalmology: current innovations, future applications and challenges.眼科领域的生成式人工智能:当前创新、未来应用和挑战。
Br J Ophthalmol. 2024 Sep 20;108(10):1335-1340. doi: 10.1136/bjo-2024-325458.

基于视觉-语言模型的光学相干断层扫描图像黄斑病变特征检测

Vision-Language Models for Feature Detection of Macular Diseases on Optical Coherence Tomography.

机构信息

Institute of Ophthalmology, University College London, London, United Kingdom.

Moorfields Eye Hospital National Health Service Foundation Trust, London, United Kingdom.

出版信息

JAMA Ophthalmol. 2024 Jun 1;142(6):573-576. doi: 10.1001/jamaophthalmol.2024.1165.

DOI:10.1001/jamaophthalmol.2024.1165
PMID:38696177
原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC11066758/
Abstract

IMPORTANCE

Vision-language models (VLMs) are a novel artificial intelligence technology capable of processing image and text inputs. While demonstrating strong generalist capabilities, their performance in ophthalmology has not been extensively studied.

OBJECTIVE

To assess the performance of the Gemini Pro VLM in expert-level tasks for macular diseases from optical coherence tomography (OCT) scans.

DESIGN, SETTING, AND PARTICIPANTS: This was a cross-sectional diagnostic accuracy study evaluating a generalist VLM on ophthalmology-specific tasks using the open-source Optical Coherence Tomography Image Database. The dataset included OCT B-scans from 50 unique patients: healthy individuals and those with macular hole, diabetic macular edema, central serous chorioretinopathy, and age-related macular degeneration. Each OCT scan was labeled for 10 key pathological features, referral recommendations, and treatments. The images were captured using a Cirrus high definition OCT machine (Carl Zeiss Meditec) at Sankara Nethralaya Eye Hospital, Chennai, India, and the dataset was published in December 2018. Image acquisition dates were not specified.

EXPOSURES

Gemini Pro, using a standard prompt to extract structured responses on December 15, 2023.

MAIN OUTCOMES AND MEASURES

The primary outcome was model responses compared against expert labels, calculating F1 scores for each pathological feature. Secondary outcomes included accuracy in diagnosis, referral urgency, and treatment recommendation. The model's internal concordance was evaluated by measuring the alignment between referral and treatment recommendations, independent of diagnostic accuracy.

RESULTS

The mean F1 score was 10.7% (95% CI, 2.4-19.2). Measurable F1 scores were obtained for macular hole (36.4%; 95% CI, 0-71.4), pigment epithelial detachment (26.1%; 95% CI, 0-46.2), subretinal hyperreflective material (24.0%; 95% CI, 0-45.2), and subretinal fluid (20.0%; 95% CI, 0-45.5). A correct diagnosis was achieved in 17 of 50 cases (34%; 95% CI, 22-48). Referral recommendations varied: 28 of 50 were correct (56%; 95% CI, 42-70), 10 of 50 were overcautious (20%; 95% CI, 10-32), and 12 of 50 were undercautious (24%; 95% CI, 12-36). Referral and treatment concordance were very high, with 48 of 50 (96%; 95 % CI, 90-100) and 48 of 49 (98%; 95% CI, 94-100) correct answers, respectively.

CONCLUSIONS AND RELEVANCE

In this study, a generalist VLM demonstrated limited vision capabilities for feature detection and management of macular disease. However, it showed low self-contradiction, suggesting strong language capabilities. As VLMs continue to improve, validating their performance on large benchmarking datasets will help ascertain their potential in ophthalmology.

摘要

重要性

视觉语言模型(VLMs)是一种新型的人工智能技术,能够处理图像和文本输入。虽然表现出很强的通才能力,但它们在眼科领域的性能尚未得到广泛研究。

目的

评估 Gemini Pro VLM 在眼科专用任务中对来自光学相干断层扫描(OCT)扫描的黄斑疾病的表现。

设计、设置和参与者:这是一项横断面诊断准确性研究,使用开源的光学相干断层扫描图像数据库评估通用 VLM 在眼科特定任务上的性能。该数据集包括来自 50 个独特患者的 OCT B 扫描:健康个体和患有黄斑裂孔、糖尿病性黄斑水肿、中心性浆液性脉络膜视网膜病变和年龄相关性黄斑变性的患者。每个 OCT 扫描都标记了 10 个关键病理特征、转诊建议和治疗方法。图像是使用 Cirrus 高清 OCT 机(卡尔蔡司 Meditec)在印度钦奈的 Sankara Nethralaya 眼科医院拍摄的,数据集于 2018 年 12 月发布。未指定图像采集日期。

暴露

2023 年 12 月 15 日,使用标准提示语提取 Gemini Pro 的结构化回复。

主要结果和措施

主要结果是将模型响应与专家标签进行比较,计算每个病理特征的 F1 分数。次要结果包括诊断准确性、转诊紧迫性和治疗建议。通过测量转诊和治疗建议之间的一致性,评估模型的内部一致性,而不考虑诊断准确性。

结果

平均 F1 分数为 10.7%(95%CI,2.4-19.2)。黄斑裂孔(36.4%;95%CI,0-71.4)、色素上皮脱离(26.1%;95%CI,0-46.2)、视网膜下高反射物质(24.0%;95%CI,0-45.2)和视网膜下液(20.0%;95%CI,0-45.5)可获得可衡量的 F1 分数。50 例中有 17 例(34%;95%CI,22-48)做出正确诊断。转诊建议各不相同:50 例中有 28 例(56%;95%CI,42-70)正确,10 例(20%;95%CI,10-32)过于谨慎,12 例(24%;95%CI,12-36)过于谨慎。转诊和治疗的一致性非常高,分别有 48 例(96%;95%CI,90-100)和 48 例(98%;95%CI,94-100)正确答案。

结论和相关性

在这项研究中,通用 VLM 对黄斑疾病的特征检测和管理表现出有限的视觉能力。然而,它表现出较低的自我矛盾,表明其具有较强的语言能力。随着 VLMs 的不断改进,在大型基准数据集上验证其性能将有助于确定它们在眼科领域的潜力。