• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

具有子组分布对齐调整的公平文本到医学图像扩散模型

Fair Text to Medical Image Diffusion Model with Subgroup Distribution Aligned Tuning.

作者信息

Han Xu, Fan Fangfang, Rong Jingzhao, Li Zhen, Fakhri Georges El, Chen Qingyu, Liu Xiaofeng

机构信息

Department of Radiology and Biomedical Imaging, Yale University, New Haven, CT 06519, USA.

Department of Neurology, Beth Israel Deaconess Medical Center, Harvard Medical School, Boston, MA 02140, USA.

出版信息

Proc SPIE Int Soc Opt Eng. 2025 Feb;13411. doi: 10.1117/12.3046450. Epub 2025 Apr 10.

DOI:10.1117/12.3046450
PMID:40831669
原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC12360154/
Abstract

The Text to Medical Image (T2MedI) approach using latent diffusion models holds significant promise for addressing the scarcity of medical imaging data and elucidating the appearance distribution of lesions corresponding to specific patient status descriptions. Like natural image synthesis models, our investigations reveal that the T2MedI model may exhibit biases towards certain subgroups, potentially neglecting minority groups present in the training dataset. In this study, we initially developed a T2MedI model adapted from the pre-trained Imagen framework. This model employs a fixed Contrastive Language-Image Pre-training (CLIP) text encoder, with its decoder fine-tuned using medical images from the Radiology Objects in Context (ROCO) dataset. We conduct both qualitative and quantitative analyses to examine its gender bias. To address this issue, we propose a subgroup distribution alignment method during fine-tuning on a target application dataset. Specifically, this process involves an alignment loss, guided by an off-the-shelf sensitivity-subgroup classifier, which aims to synchronize the classification probabilities between the generated images and those expected in the target dataset. Additionally, we preserve image quality through a CLIP-consistency regularization term, based on a knowledge distillation framework. For evaluation purposes, we designated the BraTS18 dataset as the target, and developed a gender classifier based on brain magnetic resonance (MR) imaging slices derived from it. Our methodology significantly mitigates gender representation inconsistencies in the generated MR images, aligning them more closely with the gender distribution in the BraTS18 dataset.

摘要

使用潜在扩散模型的文本到医学图像(T2MedI)方法在解决医学成像数据稀缺问题以及阐明与特定患者状态描述相对应的病变外观分布方面具有巨大潜力。与自然图像合成模型一样,我们的研究表明,T2MedI模型可能对某些亚组存在偏差,可能会忽略训练数据集中存在的少数群体。在本研究中,我们最初开发了一个基于预训练的Imagen框架改编的T2MedI模型。该模型采用固定的对比语言-图像预训练(CLIP)文本编码器,其解码器使用来自上下文放射学对象(ROCO)数据集的医学图像进行微调。我们进行了定性和定量分析以检查其性别偏差。为了解决这个问题,我们在目标应用数据集的微调过程中提出了一种亚组分布对齐方法。具体来说,这个过程涉及一个对齐损失,由一个现成的敏感性-亚组分类器引导,旨在使生成图像的分类概率与目标数据集中预期的概率同步。此外,我们基于知识蒸馏框架通过CLIP一致性正则化项来保持图像质量。为了进行评估,我们将BraTS18数据集指定为目标,并基于从中提取的脑磁共振(MR)成像切片开发了一个性别分类器。我们的方法显著减轻了生成的MR图像中性别表示的不一致性,使其与BraTS18数据集中的性别分布更紧密地对齐。

相似文献

1
Fair Text to Medical Image Diffusion Model with Subgroup Distribution Aligned Tuning.具有子组分布对齐调整的公平文本到医学图像扩散模型
Proc SPIE Int Soc Opt Eng. 2025 Feb;13411. doi: 10.1117/12.3046450. Epub 2025 Apr 10.
2
Prescription of Controlled Substances: Benefits and Risks管制药品的处方:益处与风险
3
A medical image classification method based on self-regularized adversarial learning.基于自正则化对抗学习的医学图像分类方法。
Med Phys. 2024 Nov;51(11):8232-8246. doi: 10.1002/mp.17320. Epub 2024 Jul 30.
4
Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.利用基础模型库进行跨设备肿瘤显微镜检查中的细胞相似性搜索。
Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.
5
Magnetic resonance perfusion for differentiating low-grade from high-grade gliomas at first presentation.首次就诊时磁共振灌注成像用于鉴别低级别与高级别胶质瘤
Cochrane Database Syst Rev. 2018 Jan 22;1(1):CD011551. doi: 10.1002/14651858.CD011551.pub2.
6
Exploring the Potential of Electroencephalography Signal-Based Image Generation Using Diffusion Models: Integrative Framework Combining Mixed Methods and Multimodal Analysis.利用扩散模型探索基于脑电图信号的图像生成潜力:结合混合方法和多模态分析的综合框架
JMIR Med Inform. 2025 Jun 25;13:e72027. doi: 10.2196/72027.
7
Sexual Harassment and Prevention Training性骚扰与预防培训
8
Influence of early through late fusion on pancreas segmentation from imperfectly registered multimodal magnetic resonance imaging.早期至晚期融合对来自配准不完善的多模态磁共振成像的胰腺分割的影响。
J Med Imaging (Bellingham). 2025 Mar;12(2):024008. doi: 10.1117/1.JMI.12.2.024008. Epub 2025 Apr 26.
9
Improving reconstruction of patient-specific abnormalities in AI-driven fast MRI with an individually adapted diffusion model.利用个体适配的扩散模型改进人工智能驱动的快速磁共振成像中患者特异性异常的重建。
Med Phys. 2025 Jul;52(7):e17955. doi: 10.1002/mp.17955.
10
Comparison of Two Modern Survival Prediction Tools, SORG-MLA and METSSS, in Patients With Symptomatic Long-bone Metastases Who Underwent Local Treatment With Surgery Followed by Radiotherapy and With Radiotherapy Alone.两种现代生存预测工具 SORG-MLA 和 METSSS 在接受手术联合放疗和单纯放疗治疗有症状长骨转移患者中的比较。
Clin Orthop Relat Res. 2024 Dec 1;482(12):2193-2208. doi: 10.1097/CORR.0000000000003185. Epub 2024 Jul 23.

本文引用的文献

1
Classifying sex with volume-matched brain MRI.利用体积匹配的脑部磁共振成像对性别进行分类。
Neuroimage Rep. 2023 Jul 20;3(3):100181. doi: 10.1016/j.ynirp.2023.100181. eCollection 2023 Sep.