• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

一种在小标注样本环境下基于多模态增强的半监督学习分类新方法。

A new method of semi-supervised learning classification based on multi-mode augmentation in small labeled sample environment.

作者信息

Liu Yuxuan, Wen Chenglin

机构信息

School of Automation, Guangdong University of Petrochemical Technology, Maoming, 510006, China.

出版信息

Sci Rep. 2025 Jul 1;15(1):22022. doi: 10.1038/s41598-025-02324-0.

DOI:10.1038/s41598-025-02324-0
PMID:40593789
原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC12215796/
Abstract

Semi-supervised learning mitigates the problem of labeled data scarcity by utilizing unlabeled data, but the generalization performance of existing methods usually degrades significantly when the unlabeled data is small in size or poor in quality. To this end, this paper proposes a semi-supervised image classification method based on multi-mode augmentation, which mitigates the effects of insufficient quality and limited scale of unlabeled data by simultaneously improving the sample completeness within and between classes. Specifically, the model's prediction confidence and bias are used for uncertainty-based screening to improve pseudo-label quality, while retaining as many unlabeled samples as possible to fully exploit their potential information. Secondly, a multi-modal data augmentation strategy combining intra-class random augmentation and inter-class mixed augmentation is designed to enhance the diversity of the data and the feature expression capability. Finally, a pseudo-label consistency metric is introduced to further improve the model's generalization ability. The experimental results on STL-10 and CIFAR-10 datasets show that the generalization performance of the proposed method is significantly better than the existing mainstream methods in the scenarios of small unlabeled data and mismatched samples.

摘要

半监督学习通过利用未标记数据缓解了标记数据稀缺的问题,但当未标记数据规模较小或质量较差时,现有方法的泛化性能通常会显著下降。为此,本文提出了一种基于多模态增强的半监督图像分类方法,该方法通过同时提高类内和类间的样本完整性来减轻未标记数据质量不足和规模有限的影响。具体来说,利用模型的预测置信度和偏差进行基于不确定性的筛选,以提高伪标签质量,同时保留尽可能多的未标记样本以充分挖掘其潜在信息。其次,设计了一种结合类内随机增强和类间混合增强的多模态数据增强策略,以增强数据的多样性和特征表达能力。最后,引入了伪标签一致性度量以进一步提高模型的泛化能力。在STL-10和CIFAR-10数据集上的实验结果表明,在未标记数据较少和样本不匹配的场景中,所提方法的泛化性能明显优于现有主流方法。

https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/fb4a0cebb84d/41598_2025_2324_Fig9_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/4d9097d70a4b/41598_2025_2324_Fig1_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/323f6b0ec2ff/41598_2025_2324_Fig2_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/a22573aa0621/41598_2025_2324_Fig3_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/97475ab12b92/41598_2025_2324_Fig4_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/ed57a031e345/41598_2025_2324_Fig5_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/dc1da575bcba/41598_2025_2324_Fig6_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/29223315afd7/41598_2025_2324_Fig7_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/9c6235594257/41598_2025_2324_Fig8_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/fb4a0cebb84d/41598_2025_2324_Fig9_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/4d9097d70a4b/41598_2025_2324_Fig1_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/323f6b0ec2ff/41598_2025_2324_Fig2_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/a22573aa0621/41598_2025_2324_Fig3_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/97475ab12b92/41598_2025_2324_Fig4_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/ed57a031e345/41598_2025_2324_Fig5_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/dc1da575bcba/41598_2025_2324_Fig6_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/29223315afd7/41598_2025_2324_Fig7_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/9c6235594257/41598_2025_2324_Fig8_HTML.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/0c6e/12215796/fb4a0cebb84d/41598_2025_2324_Fig9_HTML.jpg

相似文献

1
A new method of semi-supervised learning classification based on multi-mode augmentation in small labeled sample environment.一种在小标注样本环境下基于多模态增强的半监督学习分类新方法。
Sci Rep. 2025 Jul 1;15(1):22022. doi: 10.1038/s41598-025-02324-0.
2
Home treatment for mental health problems: a systematic review.心理健康问题的居家治疗:一项系统综述
Health Technol Assess. 2001;5(15):1-139. doi: 10.3310/hta5150.
3
Systemic treatments for metastatic cutaneous melanoma.转移性皮肤黑色素瘤的全身治疗
Cochrane Database Syst Rev. 2018 Feb 6;2(2):CD011123. doi: 10.1002/14651858.CD011123.pub2.
4
A rapid and systematic review of the clinical effectiveness and cost-effectiveness of paclitaxel, docetaxel, gemcitabine and vinorelbine in non-small-cell lung cancer.对紫杉醇、多西他赛、吉西他滨和长春瑞滨在非小细胞肺癌中的临床疗效和成本效益进行的快速系统评价。
Health Technol Assess. 2001;5(32):1-195. doi: 10.3310/hta5320.
5
Semantic contrast with uncertainty-aware pseudo label for lumbar semi-supervised classification.基于具有不确定性感知的伪标签的语义对比进行腰椎半监督分类。
Comput Biol Med. 2024 Aug;178:108754. doi: 10.1016/j.compbiomed.2024.108754. Epub 2024 Jun 15.
6
Signs and symptoms to determine if a patient presenting in primary care or hospital outpatient settings has COVID-19.在基层医疗机构或医院门诊环境中,如果患者出现以下症状和体征,可判断其是否患有 COVID-19。
Cochrane Database Syst Rev. 2022 May 20;5(5):CD013665. doi: 10.1002/14651858.CD013665.pub3.
7
Adefovir dipivoxil and pegylated interferon alfa-2a for the treatment of chronic hepatitis B: a systematic review and economic evaluation.阿德福韦酯与聚乙二醇化干扰素α-2a治疗慢性乙型肝炎:系统评价与经济学评估
Health Technol Assess. 2006 Aug;10(28):iii-iv, xi-xiv, 1-183. doi: 10.3310/hta10280.
8
Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.利用基础模型库进行跨设备肿瘤显微镜检查中的细胞相似性搜索。
Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.
9
Semi-Supervised Learning Allows for Improved Segmentation With Reduced Annotations of Brain Metastases Using Multicenter MRI Data.半监督学习可利用多中心MRI数据,通过减少脑转移瘤的标注来改进分割。
J Magn Reson Imaging. 2025 Jun;61(6):2469-2479. doi: 10.1002/jmri.29686. Epub 2025 Jan 10.
10
Systemic pharmacological treatments for chronic plaque psoriasis: a network meta-analysis.系统性药理学治疗慢性斑块状银屑病:网络荟萃分析。
Cochrane Database Syst Rev. 2021 Apr 19;4(4):CD011535. doi: 10.1002/14651858.CD011535.pub4.

本文引用的文献

1
Multilabel Feature Selection via Shared Latent Sublabel Structure and Simultaneous Orthogonal Basis Clustering.基于共享潜在子标签结构和同步正交基聚类的多标签特征选择
IEEE Trans Neural Netw Learn Syst. 2025 Mar;36(3):5288-5303. doi: 10.1109/TNNLS.2024.3382911. Epub 2025 Feb 28.
2
Semi-Supervised Disease Classification Based on Limited Medical Image Data.基于有限医学图像数据的半监督疾病分类。
IEEE J Biomed Health Inform. 2024 Mar;28(3):1575-1586. doi: 10.1109/JBHI.2024.3349412. Epub 2024 Mar 6.
3
Interpolation consistency training for semi-supervised learning.
半监督学习中的插值一致性训练。
Neural Netw. 2022 Jan;145:90-106. doi: 10.1016/j.neunet.2021.10.008. Epub 2021 Oct 21.
4
FMixCutMatch for semi-supervised deep learning.FMixCutMatch 用于半监督深度学习。
Neural Netw. 2021 Jan;133:166-176. doi: 10.1016/j.neunet.2020.10.018. Epub 2020 Nov 10.
5
Biomedical image classification made easier thanks to transfer and semi-supervised learning.得益于迁移学习和半监督学习,生物医学图像分类变得更加容易。
Comput Methods Programs Biomed. 2021 Jan;198:105782. doi: 10.1016/j.cmpb.2020.105782. Epub 2020 Oct 3.