• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

基于拉曼光谱的XGBoost算法辅助多组分定量分析

XGBoost algorithm assisted multi-component quantitative analysis with Raman spectroscopy.

作者信息

Wang Qiaoyun, Zou Xin, Chen Yinji, Zhu Ziheng, Yan Chongyue, Shan Peng, Wang Shuyu, Fu Yongqing

机构信息

College of Information Science and Engineering, Northeastern University, Shenyang, Liaoning Province 110819, China; Hebei Key Laboratory of Micro-Nano Precision Optical Sensing and Measurement Technology, Qinhuangdao 066004, China.

College of Information Science and Engineering, Northeastern University, Shenyang, Liaoning Province 110819, China.

出版信息

Spectrochim Acta A Mol Biomol Spectrosc. 2024 Dec 15;323:124917. doi: 10.1016/j.saa.2024.124917. Epub 2024 Jul 31.

DOI:10.1016/j.saa.2024.124917
PMID:39094267
Abstract

To improve prediction performance and reduce artifacts in Raman spectra, we developed an eXtreme Gradient Boosting (XGBoost) preprocessing method to preprocess the Raman spectra of glucose, glycerol and ethanol mixtures. To ensure the robustness and reliability of the XGBoost preprocessing method, an X-LR model (which combined XGBoost preprocessing and a linear regression (LR) model) and a X-MLP model (which combined XGBoost preprocessing and a multilayer perceptron (MLP) model) were developed. These two models were used to quantitatively analyze the concentrations of glucose, glycerol and ethanol in the Raman spectra of mixed solutions. The proportion map of hyperparameters was firstly used to narrow down the search range of hyperparameters in the X-LR and the X-MLP models. Then the correlation coefficients (R), root mean square of calibration (RMSEC), and root mean square error of prediction (RMSEP) were used to evaluate the models' performance. Experimental results indicated that the XGBoost preprocessing method achieved higher accuracy and generalization capability, with better performance than those of other preprocessing methods for both LR and MLP models.

摘要

为了提高拉曼光谱的预测性能并减少伪影,我们开发了一种极端梯度提升(XGBoost)预处理方法,用于对葡萄糖、甘油和乙醇混合物的拉曼光谱进行预处理。为确保XGBoost预处理方法的稳健性和可靠性,我们开发了一个X-LR模型(结合了XGBoost预处理和线性回归(LR)模型)和一个X-MLP模型(结合了XGBoost预处理和多层感知器(MLP)模型)。这两个模型用于定量分析混合溶液拉曼光谱中葡萄糖、甘油和乙醇的浓度。首先使用超参数比例图来缩小X-LR和X-MLP模型中超参数的搜索范围。然后使用相关系数(R)、校准均方根(RMSEC)和预测均方根误差(RMSEP)来评估模型的性能。实验结果表明,XGBoost预处理方法具有更高的准确性和泛化能力,对于LR和MLP模型,其性能均优于其他预处理方法。

相似文献

1
XGBoost algorithm assisted multi-component quantitative analysis with Raman spectroscopy.基于拉曼光谱的XGBoost算法辅助多组分定量分析
Spectrochim Acta A Mol Biomol Spectrosc. 2024 Dec 15;323:124917. doi: 10.1016/j.saa.2024.124917. Epub 2024 Jul 31.
2
Predicting the Tool Wear of a Drilling Process Using Novel Machine Learning XGBoost-SDA.使用新型机器学习XGBoost-SDA预测钻孔过程中的刀具磨损
Materials (Basel). 2020 Nov 4;13(21):4952. doi: 10.3390/ma13214952.
3
Chemometric Analysis of a Ternary Mixture of Caffeine, Quinic Acid, and Nicotinic Acid by Terahertz Spectroscopy.基于太赫兹光谱的咖啡因、奎尼酸和烟酸三元混合物的化学计量学分析
ACS Omega. 2022 Sep 27;7(40):35783-35791. doi: 10.1021/acsomega.2c03808. eCollection 2022 Oct 11.
4
Raman spectroscopy combined with partial least squares (PLS) based on hybrid spectral preprocessing and backward interval PLS (biPLS) for quantitative analysis of four PAHs in oil sludge.基于混合光谱预处理和后向区间偏最小二乘法(biPLS)的拉曼光谱结合偏最小二乘法(PLS)用于油泥中四种多环芳烃的定量分析。
Spectrochim Acta A Mol Biomol Spectrosc. 2024 Apr 5;310:123953. doi: 10.1016/j.saa.2024.123953. Epub 2024 Jan 24.
5
Application of RR-XGBoost combined model in data calibration of micro air quality detector.RR-XGBoost 组合模型在微空气质量检测仪数据校准中的应用。
Sci Rep. 2021 Aug 2;11(1):15662. doi: 10.1038/s41598-021-95027-1.
6
A machine learning model based on ultrasound image features to assess the risk of sentinel lymph node metastasis in breast cancer patients: Applications of scikit-learn and SHAP.一种基于超声图像特征的机器学习模型,用于评估乳腺癌患者前哨淋巴结转移风险:scikit-learn和SHAP的应用
Front Oncol. 2022 Jul 25;12:944569. doi: 10.3389/fonc.2022.944569. eCollection 2022.
7
On the Use of Machine Learning Models for Prediction of Compressive Strength of Concrete: Influence of Dimensionality Reduction on the Model Performance.关于使用机器学习模型预测混凝土抗压强度:降维对模型性能的影响
Materials (Basel). 2021 Feb 3;14(4):713. doi: 10.3390/ma14040713.
8
A Retrospective Cohort Study: Predicting 90-Day Mortality for ICU Trauma Patients with a Machine Learning Algorithm Using XGBoost Using MIMIC-III Database.一项回顾性队列研究:使用MIMIC-III数据库,通过基于XGBoost的机器学习算法预测ICU创伤患者的90天死亡率。
J Multidiscip Healthc. 2023 Sep 6;16:2625-2640. doi: 10.2147/JMDH.S416943. eCollection 2023.
9
A Hybrid Model for Temperature Prediction in a Sheep House.羊舍温度预测的混合模型
Animals (Basel). 2022 Oct 17;12(20):2806. doi: 10.3390/ani12202806.
10
Identification of High-Risk Patients for Postoperative Myocardial Injury After CME Using Machine Learning: A 10-Year Multicenter Retrospective Study.使用机器学习识别结直肠癌根治术后心肌损伤的高危患者:一项为期10年的多中心回顾性研究
Int J Gen Med. 2023 Apr 7;16:1251-1264. doi: 10.2147/IJGM.S409363. eCollection 2023.

引用本文的文献

1
Study on the quantitative analysis of Tilianin based on Raman spectroscopy combined with deep learning.基于拉曼光谱结合深度学习的田基黄宁定量分析研究
PLoS One. 2025 Jun 18;20(6):e0325530. doi: 10.1371/journal.pone.0325530. eCollection 2025.
2
Optimizing hypoglycaemia prediction in type 1 diabetes with Ensemble Machine Learning modeling.使用集成机器学习模型优化1型糖尿病患者的低血糖预测
BMC Med Inform Decis Mak. 2025 Jan 31;25(1):46. doi: 10.1186/s12911-025-02867-2.