• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

TGFuse:一种基于Transformer和生成对抗网络的红外与可见光图像融合方法

TGFuse: An Infrared and Visible Image Fusion Approach Based on Transformer and Generative Adversarial Network.

作者信息

Rao Dongyu, Xu Tianyang, Wu Xiao-Jun

出版信息

IEEE Trans Image Process. 2023 May 10;PP. doi: 10.1109/TIP.2023.3273451.

DOI:10.1109/TIP.2023.3273451
PMID:37163395
Abstract

The end-to-end image fusion framework has achieved promising performance, with dedicated convolutional networks aggregating the multi-modal local appearance. However, long-range dependencies are directly neglected in existing CNN fusion approaches, impeding balancing the entire image-level perception for complex scenario fusion. In this paper, therefore, we propose an infrared and visible image fusion algorithm based on the transformer module and adversarial learning. Inspired by the global interaction power, we use the transformer technique to learn the effective global fusion relations. In particular, shallow features extracted by CNN are interacted in the proposed transformer fusion module to refine the fusion relationship within the spatial scope and across channels simultaneously. Besides, adversarial learning is designed in the training process to improve the output discrimination via imposing competitive consistency from the inputs, reflecting the specific characteristics in infrared and visible images. The experimental performance demonstrates the effectiveness of the proposed modules, with superior improvement against the state-of-the-art, generalising a novel paradigm via transformer and adversarial learning in the fusion task.

摘要

端到端图像融合框架已经取得了不错的性能,通过专用卷积网络聚合多模态局部外观。然而,现有卷积神经网络(CNN)融合方法直接忽略了长距离依赖关系,这阻碍了在复杂场景融合中平衡整个图像级别的感知。因此,在本文中,我们提出了一种基于Transformer模块和对抗学习的红外与可见光图像融合算法。受全局交互能力的启发,我们使用Transformer技术来学习有效的全局融合关系。具体而言,由CNN提取的浅层特征在所提出的Transformer融合模块中进行交互,以同时在空间范围内和跨通道细化融合关系。此外,在训练过程中设计了对抗学习,通过从输入施加竞争一致性来提高输出的判别力,反映红外和可见光图像中的特定特征。实验性能证明了所提出模块的有效性,相对于现有技术有显著改进,在融合任务中通过Transformer和对抗学习推广了一种新的范式。

相似文献

1
TGFuse: An Infrared and Visible Image Fusion Approach Based on Transformer and Generative Adversarial Network.TGFuse:一种基于Transformer和生成对抗网络的红外与可见光图像融合方法
IEEE Trans Image Process. 2023 May 10;PP. doi: 10.1109/TIP.2023.3273451.
2
CTFusion: CNN-transformer-based self-supervised learning for infrared and visible image fusion.CTFusion:基于卷积神经网络-Transformer的红外与可见光图像融合自监督学习
Math Biosci Eng. 2024 Jul 30;21(7):6710-6730. doi: 10.3934/mbe.2024294.
3
Dual encoder network with transformer-CNN for multi-organ segmentation.基于 Transformer-CNN 的双编码器网络的多器官分割。
Med Biol Eng Comput. 2023 Mar;61(3):661-671. doi: 10.1007/s11517-022-02723-9. Epub 2022 Dec 29.
4
Automated multi-modal Transformer network (AMTNet) for 3D medical images segmentation.用于3D医学图像分割的自动多模态Transformer网络(AMTNet)。
Phys Med Biol. 2023 Jan 9;68(2). doi: 10.1088/1361-6560/aca74c.
5
Infrared-Visible Image Fusion Based on Semantic Guidance and Visual Perception.基于语义引导和视觉感知的红外-可见光图像融合
Entropy (Basel). 2022 Sep 21;24(10):1327. doi: 10.3390/e24101327.
6
Microscopic Hyperspectral Image Classification Based on Fusion Transformer With Parallel CNN.基于融合 Transformer 与并行 CNN 的微观高光谱图像分类。
IEEE J Biomed Health Inform. 2023 Jun;27(6):2910-2921. doi: 10.1109/JBHI.2023.3253722. Epub 2023 Jun 5.
7
Fusion Networks of CNN and Transformer with Channel Attention for Accurate Tumor Imaging in Magnetic Particle Imaging.具有通道注意力机制的CNN与Transformer融合网络用于磁粒子成像中的精确肿瘤成像
Biology (Basel). 2023 Dec 19;13(1):0. doi: 10.3390/biology13010002.
8
DSA-Net: Infrared and Visible Image Fusion via Dual-Stream Asymmetric Network.DSA-Net:通过双流非对称网络实现红外与可见光图像融合
Sensors (Basel). 2023 Aug 11;23(16):7097. doi: 10.3390/s23167097.
9
A transformer-based generative adversarial network for brain tumor segmentation.一种用于脑肿瘤分割的基于Transformer的生成对抗网络。
Front Neurosci. 2022 Nov 30;16:1054948. doi: 10.3389/fnins.2022.1054948. eCollection 2022.
10
Visible-Image-Assisted Nonuniformity Correction of Infrared Images Using the GAN with SEBlock.使用带 SEBlock 的 GAN 进行可见图像辅助的红外图像非均匀性校正。
Sensors (Basel). 2023 Mar 20;23(6):3282. doi: 10.3390/s23063282.

引用本文的文献

1
Saliency-enhanced infrared and visible image fusion via sub-window variance filter and weighted least squares optimization.基于子窗口方差滤波器和加权最小二乘优化的显著性增强红外与可见光图像融合
PLoS One. 2025 Jul 7;20(7):e0323285. doi: 10.1371/journal.pone.0323285. eCollection 2025.
2
SC-CoSF: Self-Correcting Collaborative and Co-Training for Image Fusion and Semantic Segmentation.SC-CoSF:用于图像融合与语义分割的自校正协作与协同训练
Sensors (Basel). 2025 Jun 6;25(12):3575. doi: 10.3390/s25123575.
3
Robust Infrared-Visible Fusion Imaging with Decoupled Semantic Segmentation Network.
基于解耦语义分割网络的稳健红外-可见光融合成像
Sensors (Basel). 2025 Apr 22;25(9):2646. doi: 10.3390/s25092646.
4
A stochastic structural similarity guided approach for multi-modal medical image fusion.一种基于随机结构相似性引导的多模态医学图像融合方法。
Sci Rep. 2025 Mar 14;15(1):8792. doi: 10.1038/s41598-025-93662-6.
5
Semantic-Aware Fusion Network Based on Super-Resolution.基于超分辨率的语义感知融合网络
Sensors (Basel). 2024 Jun 5;24(11):3665. doi: 10.3390/s24113665.
6
TGLFusion: A Temperature-Guided Lightweight Fusion Method for Infrared and Visible Images.TGLFusion:一种用于红外与可见光图像的温度引导轻量级融合方法
Sensors (Basel). 2024 Mar 7;24(6):1735. doi: 10.3390/s24061735.
7
Using Sparse Parts in Fused Information to Enhance Performance in Latent Low-Rank Representation-Based Fusion of Visible and Infrared Images.在融合信息中使用稀疏部分以增强基于潜在低秩表示的可见光与红外图像融合性能
Sensors (Basel). 2024 Feb 26;24(5):1514. doi: 10.3390/s24051514.
8
SFPFusion: An Improved Vision Transformer Combining Super Feature Attention and Wavelet-Guided Pooling for Infrared and Visible Images Fusion.SFPFusion:一种改进的视觉Transformer,结合超级特征注意力和小波引导池化用于红外与可见光图像融合。
Sensors (Basel). 2023 Sep 13;23(18):7870. doi: 10.3390/s23187870.
9
DPACFuse: Dual-Branch Progressive Learning for Infrared and Visible Image Fusion with Complementary Self-Attention and Convolution.DPACFuse:基于互补自注意力与卷积的红外与可见光图像融合双分支渐进学习
Sensors (Basel). 2023 Aug 16;23(16):7205. doi: 10.3390/s23167205.
10
DSA-Net: Infrared and Visible Image Fusion via Dual-Stream Asymmetric Network.DSA-Net:通过双流非对称网络实现红外与可见光图像融合
Sensors (Basel). 2023 Aug 11;23(16):7097. doi: 10.3390/s23167097.