• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

基于对比标签注意力的多标签零样本学习

Multi-Label Zero-Shot Learning Via Contrastive Label-Based Attention.

作者信息

Meng Shixuan, Jiang Rongxin, Tian Xiang, Zhou Fan, Chen Yaowu, Liu Junjie, Shen Chen

机构信息

Zhejiang University, Hangzhou, P. R. China.

Zhejiang University, Embedded System Engineering Research Center, Ministry of Education of China, Hangzhou, P. R. China.

出版信息

Int J Neural Syst. 2025 Mar;35(3):2550010. doi: 10.1142/S0129065725500108. Epub 2025 Jan 23.

DOI:10.1142/S0129065725500108
PMID:39851003
Abstract

Multi-label zero-shot learning (ML-ZSL) strives to recognize all objects in an image, regardless of whether they are present in the training data. Recent methods incorporate an attention mechanism to locate labels in the image and generate class-specific semantic information. However, the attention mechanism built on visual features treats label embeddings equally in the prediction score, leading to severe semantic ambiguity. This study focuses on efficiently utilizing semantic information in the attention mechanism. We propose a contrastive label-based attention method (CLA) to associate each label with the most relevant image regions. Specifically, our label-based attention, guided by the latent label embedding, captures discriminative image details. To distinguish region-wise correlations, we implement a region-level contrastive loss. In addition, we utilize a global feature alignment module to identify labels with general information. Extensive experiments on two benchmarks, NUS-WIDE and Open Images, demonstrate that our CLA outperforms the state-of-the-art methods. Especially under the ZSL setting, our method achieves 2.0% improvements in mean Average Precision (mAP) for NUS-WIDE and 4.0% for Open Images compared with recent methods.

摘要

多标签零样本学习(ML-ZSL)致力于识别图像中的所有物体,无论它们是否出现在训练数据中。最近的方法引入了注意力机制来定位图像中的标签并生成特定类别的语义信息。然而,基于视觉特征构建的注意力机制在预测分数中对标签嵌入一视同仁,导致严重的语义模糊。本研究专注于在注意力机制中有效利用语义信息。我们提出了一种基于对比标签的注意力方法(CLA),将每个标签与最相关的图像区域相关联。具体而言,我们基于标签的注意力在潜在标签嵌入的引导下,捕捉有区分力的图像细节。为了区分区域级别的相关性,我们实现了一种区域级对比损失。此外,我们利用一个全局特征对齐模块来识别具有一般信息的标签。在NUS-WIDE和Open Images这两个基准上进行的大量实验表明,我们的CLA优于当前的先进方法。特别是在零样本学习设置下,与最近的方法相比,我们的方法在NUS-WIDE上的平均精度均值(mAP)提高了2.0%,在Open Images上提高了4.0%。

相似文献

1
Multi-Label Zero-Shot Learning Via Contrastive Label-Based Attention.基于对比标签注意力的多标签零样本学习
Int J Neural Syst. 2025 Mar;35(3):2550010. doi: 10.1142/S0129065725500108. Epub 2025 Jan 23.
2
Graph embedding based multi-label Zero-shot Learning.基于图嵌入的多标签零样本学习。
Neural Netw. 2023 Oct;167:129-140. doi: 10.1016/j.neunet.2023.08.023. Epub 2023 Aug 19.
3
Multi-label zero-shot learning with graph convolutional networks.基于图卷积网络的多标签零样本学习。
Neural Netw. 2020 Dec;132:333-341. doi: 10.1016/j.neunet.2020.09.010. Epub 2020 Sep 21.
4
DRTN: Dual Relation Transformer Network with feature erasure and contrastive learning for multi-label image classification.DRTN:用于多标签图像分类的具有特征擦除和对比学习的双关系Transformer网络。
Neural Netw. 2025 Jul;187:107309. doi: 10.1016/j.neunet.2025.107309. Epub 2025 Mar 3.
5
Visual-guided attentive attributes embedding for zero-shot learning.基于视觉引导的注意力属性嵌入的零样本学习。
Neural Netw. 2021 Nov;143:709-718. doi: 10.1016/j.neunet.2021.07.031. Epub 2021 Aug 11.
6
Augmented semantic feature based generative network for generalized zero-shot learning.基于增强语义特征的生成网络用于广义零样本学习。
Neural Netw. 2021 Nov;143:1-11. doi: 10.1016/j.neunet.2021.04.014. Epub 2021 Apr 21.
7
Multi-label zero-shot human action recognition via joint latent ranking embedding.基于联合潜在排序嵌入的多标签零镜头人体动作识别。
Neural Netw. 2020 Feb;122:1-23. doi: 10.1016/j.neunet.2019.09.029. Epub 2019 Oct 21.
8
Label-activating framework for zero-shot learning.标签激活框架用于零样本学习。
Neural Netw. 2020 Jan;121:1-9. doi: 10.1016/j.neunet.2019.08.023. Epub 2019 Sep 6.
9
Generative Multi-Label Zero-Shot Learning.生成式多标签零样本学习
IEEE Trans Pattern Anal Mach Intell. 2023 Dec;45(12):14611-14624. doi: 10.1109/TPAMI.2023.3295772. Epub 2023 Nov 3.
10
GNDAN: Graph Navigated Dual Attention Network for Zero-Shot Learning.GNDAN:用于零样本学习的图导航双注意力网络。
IEEE Trans Neural Netw Learn Syst. 2024 Apr;35(4):4516-4529. doi: 10.1109/TNNLS.2022.3155602. Epub 2024 Apr 4.