• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

HyperSIGMA:高光谱智能理解基础模型。

HyperSIGMA: Hyperspectral Intelligence Comprehension Foundation Model.

作者信息

Wang Di, Hu Meiqi, Jin Yao, Miao Yuchun, Yang Jiaqi, Xu Yichu, Qin Xiaolei, Ma Jiaqi, Sun Lingyu, Li Chenxing, Fu Chuan, Chen Hongruixuan, Han Chengxi, Yokoya Naoto, Zhang Jing, Xu Minqiang, Liu Lin, Zhang Lefei, Wu Chen, Du Bo, Tao Dacheng, Zhang Liangpei

出版信息

IEEE Trans Pattern Anal Mach Intell. 2025 Aug;47(8):6427-6444. doi: 10.1109/TPAMI.2025.3557581.

DOI:10.1109/TPAMI.2025.3557581
PMID:40198301
Abstract

Accurate hyperspectral image (HSI) interpretation is critical for providing valuable insights into various earth observation-related applications such as urban planning, precision agriculture, and environmental monitoring. However, existing HSI processing methods are predominantly task-specific and scene-dependent, which severely limits their ability to transfer knowledge across tasks and scenes, thereby reducing the practicality in real-world applications. To address these challenges, we present HyperSIGMA, a vision transformer-based foundation model that unifies HSI interpretation across tasks and scenes, scalable to over one billion parameters. To overcome the spectral and spatial redundancy inherent in HSIs, we introduce a novel sparse sampling attention (SSA) mechanism, which effectively promotes the learning of diverse contextual features and serves as the basic block of HyperSIGMA. HyperSIGMA integrates spatial and spectral features using a specially designed spectral enhancement module. In addition, we construct a large-scale hyperspectral dataset, HyperGlobal-450K, for pre-training, which contains about 450 K hyperspectral images, significantly surpassing existing datasets in scale. Extensive experiments on various high-level and low-level HSI tasks demonstrate HyperSIGMA's versatility and superior representational capability compared to current state-of-the-art methods. Moreover, HyperSIGMA shows significant advantages in scalability, robustness, cross-modal transferring capability, real-world applicability, and computational efficiency.

摘要

准确的高光谱图像(HSI)解释对于深入了解各种与地球观测相关的应用(如城市规划、精准农业和环境监测)至关重要。然而,现有的HSI处理方法主要是针对特定任务和场景的,这严重限制了它们跨任务和场景转移知识的能力,从而降低了在实际应用中的实用性。为了应对这些挑战,我们提出了HyperSIGMA,这是一种基于视觉Transformer的基础模型,它统一了跨任务和场景的HSI解释,可扩展到超过10亿个参数。为了克服HSIs中固有的光谱和空间冗余,我们引入了一种新颖的稀疏采样注意力(SSA)机制,该机制有效地促进了对多样上下文特征的学习,并作为HyperSIGMA的基本模块。HyperSIGMA使用专门设计的光谱增强模块集成空间和光谱特征。此外,我们构建了一个大规模的高光谱数据集HyperGlobal-450K用于预训练,其中包含约450K幅高光谱图像,在规模上显著超过现有数据集。在各种高级和低级HSI任务上进行的大量实验表明,与当前最先进的方法相比,HyperSIGMA具有通用性和卓越的表征能力。此外,HyperSIGMA在可扩展性、鲁棒性、跨模态转移能力、实际适用性和计算效率方面显示出显著优势。

相似文献

1
HyperSIGMA: Hyperspectral Intelligence Comprehension Foundation Model.HyperSIGMA:高光谱智能理解基础模型。
IEEE Trans Pattern Anal Mach Intell. 2025 Aug;47(8):6427-6444. doi: 10.1109/TPAMI.2025.3557581.
2
Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.利用基础模型库进行跨设备肿瘤显微镜检查中的细胞相似性搜索。
Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.
3
Sparse-view spectral CT reconstruction via a coupled subspace representation and score-based generative model.基于耦合子空间表示和基于分数的生成模型的稀疏视图光谱CT重建
Quant Imaging Med Surg. 2025 Jun 6;15(6):5474-5495. doi: 10.21037/qims-24-2226. Epub 2025 May 28.
4
DASNet a dual branch multi level attention sheep counting network.DASNet是一种双分支多级注意力羊只计数网络。
Sci Rep. 2025 Jul 2;15(1):23228. doi: 10.1038/s41598-025-97929-w.
5
CDFAN: Cross-Domain Fusion Attention Network for Pansharpening.CDFAN:用于图像锐化的跨域融合注意力网络。
Entropy (Basel). 2025 May 27;27(6):567. doi: 10.3390/e27060567.
6
Advancing respiratory disease diagnosis: A deep learning and vision transformer-based approach with a novel X-ray dataset.推进呼吸系统疾病诊断:一种基于深度学习和视觉Transformer的方法及新型X射线数据集
Comput Biol Med. 2025 Aug;194:110501. doi: 10.1016/j.compbiomed.2025.110501. Epub 2025 Jun 9.
7
Decoding Spatial Tissue Architecture: A Scalable Bayesian Topic Model for Multiplexed Imaging Analysis.解码空间组织结构:一种用于多重成像分析的可扩展贝叶斯主题模型
bioRxiv. 2024 Nov 9:2024.10.08.617293. doi: 10.1101/2024.10.08.617293.
8
Wood Waste Valorization and Classification Approaches: A systematic review.木材废料的增值与分类方法:一项系统综述
Open Res Eur. 2025 May 6;5:5. doi: 10.12688/openreseurope.18862.1. eCollection 2025.
9
A fake news detection model using the integration of multimodal attention mechanism and residual convolutional network.一种融合多模态注意力机制和残差卷积网络的假新闻检测模型。
Sci Rep. 2025 Jul 1;15(1):20544. doi: 10.1038/s41598-025-05702-w.
10
Classification of maize seed hyperspectral images based on variable-depth convolutional kernels.基于可变深度卷积核的玉米种子高光谱图像分类
Front Plant Sci. 2025 Jun 6;16:1599231. doi: 10.3389/fpls.2025.1599231. eCollection 2025.

引用本文的文献

1
From spectrum to yield: advances in crop photosynthesis with hyperspectral imaging.从光谱到产量:利用高光谱成像技术实现作物光合作用的进展
Photosynthetica. 2025 Jul 8;63(2):196-233. doi: 10.32615/ps.2025.012. eCollection 2025.