• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

用于异常检测的基于特征的正态模型

Feature-Based Normality Models for Anomaly Detection.

作者信息

Teh Hui Yie, Wang Kevin I-Kai, Kempa-Liehr Andreas W

机构信息

Department of Electrical, Computer and Software Engineering, The University of Auckland, Auckland 1142, New Zealand.

Department of Engineering Science and Biomedical Engineering, The University of Auckland, Auckland 1142, New Zealand.

出版信息

Sensors (Basel). 2025 Aug 1;25(15):4757. doi: 10.3390/s25154757.

DOI:10.3390/s25154757
PMID:40807922
原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC12349055/
Abstract

Detecting previously unseen anomalies in sensor data is a challenging problem for artificial intelligence when sensor-specific and deployment-specific characteristics of the time series need to be learned from a short calibration period. From the application point of view, this challenge becomes increasingly important because many applications are gravitating towards utilising low-cost sensors for Internet of Things deployments. While these sensors offer cost-effectiveness and customisation, their data quality does not match that of their high-end counterparts. To improve sensor data quality while addressing the challenges of anomaly detection in Internet of Things applications, we present an anomaly detection framework that learns a normality model of sensor data. The framework models the typical behaviour of individual sensors, which is crucial for the reliable detection of sensor data anomalies, especially when dealing with sensors observing significantly different signal characteristics. Our framework learns sensor-specific normality models from a small set of anomaly-free training data while employing an unsupervised feature engineering approach to select statistically significant features. The selected features are subsequently used to train a Local Outlier Factor anomaly detection model, which adaptively determines the boundary separating normal data from anomalies. The proposed anomaly detection framework is evaluated on three real-world public environmental monitoring datasets with heterogeneous sensor readings. The sensor-specific normality models are learned from extremely short calibration periods (as short as the first 3 days or 10% of the total recorded data) and outperform four other state-of-the-art anomaly detection approaches with respect to F1-score (between 5.4% and 9.3% better) and Matthews correlation coefficient (between 4.0% and 7.6% better).

摘要

当需要从较短的校准周期中学习时间序列的传感器特定和部署特定特征时,检测传感器数据中先前未见过的异常对于人工智能来说是一个具有挑战性的问题。从应用的角度来看,这一挑战变得越来越重要,因为许多应用正倾向于在物联网部署中使用低成本传感器。虽然这些传感器具有成本效益和可定制性,但其数据质量与高端传感器不匹配。为了提高传感器数据质量并应对物联网应用中的异常检测挑战,我们提出了一个异常检测框架,该框架可以学习传感器数据的正常模型。该框架对单个传感器的典型行为进行建模,这对于可靠检测传感器数据异常至关重要,尤其是在处理观测到显著不同信号特征的传感器时。我们的框架从一小部分无异常的训练数据中学习传感器特定的正常模型,同时采用无监督特征工程方法来选择具有统计显著性的特征。随后,所选特征用于训练局部离群因子异常检测模型,该模型自适应地确定将正常数据与异常数据分开的边界。我们在三个具有异构传感器读数的真实世界公共环境监测数据集上对提出的异常检测框架进行了评估。传感器特定的正常模型是从极短的校准周期(短至前3天或总记录数据的10%)中学习得到的,在F1分数(提高5.4%至9.3%)和马修斯相关系数(提高4.0%至7.6%)方面优于其他四种先进的异常检测方法。

https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/ef3acf6ded15/sensors-25-04757-g006.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/caa4d3f57d53/sensors-25-04757-g001.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/a33f2818b232/sensors-25-04757-g002.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/4ad96212af3c/sensors-25-04757-g003.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/8f2ee038e1a3/sensors-25-04757-g004.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/6053a047e403/sensors-25-04757-g005.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/ef3acf6ded15/sensors-25-04757-g006.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/caa4d3f57d53/sensors-25-04757-g001.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/a33f2818b232/sensors-25-04757-g002.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/4ad96212af3c/sensors-25-04757-g003.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/8f2ee038e1a3/sensors-25-04757-g004.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/6053a047e403/sensors-25-04757-g005.jpg
https://cdn.ncbi.nlm.nih.gov/pmc/blobs/c3e7/12349055/ef3acf6ded15/sensors-25-04757-g006.jpg

相似文献

1
Feature-Based Normality Models for Anomaly Detection.用于异常检测的基于特征的正态模型
Sensors (Basel). 2025 Aug 1;25(15):4757. doi: 10.3390/s25154757.
2
Prescription of Controlled Substances: Benefits and Risks管制药品的处方:益处与风险
3
SentemQC - A novel and cost-efficient method for quality assurance and quality control of high-resolution frequency sensor data in fresh waters.SentemQC——一种用于淡水高分辨率频率传感器数据质量保证和质量控制的新颖且经济高效的方法。
Open Res Eur. 2024 Nov 7;4:244. doi: 10.12688/openreseurope.18134.1. eCollection 2024.
4
Does the Presence of Missing Data Affect the Performance of the SORG Machine-learning Algorithm for Patients With Spinal Metastasis? Development of an Internet Application Algorithm.缺失数据的存在是否会影响 SORG 机器学习算法在脊柱转移瘤患者中的性能?开发一种互联网应用算法。
Clin Orthop Relat Res. 2024 Jan 1;482(1):143-157. doi: 10.1097/CORR.0000000000002706. Epub 2023 Jun 12.
5
Short-Term Memory Impairment短期记忆障碍
6
Sexual Harassment and Prevention Training性骚扰与预防培训
7
Proof of concept of a fully unsupervised anomaly detection framework in CBCT-guided radiotherapy.CBCT引导放射治疗中完全无监督异常检测框架的概念验证
Med Phys. 2025 Aug;52(8):e18020. doi: 10.1002/mp.18020.
8
A rapid and systematic review of the clinical effectiveness and cost-effectiveness of paclitaxel, docetaxel, gemcitabine and vinorelbine in non-small-cell lung cancer.对紫杉醇、多西他赛、吉西他滨和长春瑞滨在非小细胞肺癌中的临床疗效和成本效益进行的快速系统评价。
Health Technol Assess. 2001;5(32):1-195. doi: 10.3310/hta5320.
9
Comparison of Two Modern Survival Prediction Tools, SORG-MLA and METSSS, in Patients With Symptomatic Long-bone Metastases Who Underwent Local Treatment With Surgery Followed by Radiotherapy and With Radiotherapy Alone.两种现代生存预测工具 SORG-MLA 和 METSSS 在接受手术联合放疗和单纯放疗治疗有症状长骨转移患者中的比较。
Clin Orthop Relat Res. 2024 Dec 1;482(12):2193-2208. doi: 10.1097/CORR.0000000000003185. Epub 2024 Jul 23.
10
Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.利用基础模型库进行跨设备肿瘤显微镜检查中的细胞相似性搜索。
Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.

本文引用的文献

1
Low-Cost Water Quality Sensors for IoT: A Systematic Review.低成本物联网水质传感器:系统评价。
Sensors (Basel). 2023 Apr 30;23(9):4424. doi: 10.3390/s23094424.
2
Feasibility of low-cost particle sensor types in long-term indoor air pollution health studies after repeated calibration, 2019-2021.2019-2021 年,反复校准后低成本粒子传感器类型在长期室内空气污染健康研究中的可行性。
Sci Rep. 2022 Aug 26;12(1):14571. doi: 10.1038/s41598-022-18200-0.
3
Toward Integrated Large-Scale Environmental Monitoring Using WSN/UAV/Crowdsensing: A Review of Applications, Signal Processing, and Future Perspectives.
利用 WSN/UAV/Crowdsensing 实现综合大规模环境监测:应用、信号处理及未来展望综述。
Sensors (Basel). 2022 Feb 25;22(5):1824. doi: 10.3390/s22051824.
4
Long-term evaluation of a low-cost air sensor network for monitoring indoor and outdoor air quality at the community scale.长期评估低成本空气传感器网络在社区范围内监测室内和室外空气质量的效果。
Sci Total Environ. 2022 Feb 10;807(Pt 2):150797. doi: 10.1016/j.scitotenv.2021.150797. Epub 2021 Oct 6.
5
The advantages of the Matthews correlation coefficient (MCC) over F1 score and accuracy in binary classification evaluation.马修斯相关系数(MCC)在二分类评估中优于 F1 得分和准确率的优势。
BMC Genomics. 2020 Jan 2;21(1):6. doi: 10.1186/s12864-019-6413-7.
6
Normality: Part descriptive, part prescriptive.常态:部分是描述性的,部分是规范性的。
Cognition. 2017 Oct;167:25-37. doi: 10.1016/j.cognition.2016.10.024. Epub 2016 Nov 11.
7
Estimating the support of a high-dimensional distribution.估计高维分布的支撑集。
Neural Comput. 2001 Jul;13(7):1443-71. doi: 10.1162/089976601750264965.