基于化学信息和靶标活性导向算法的最优描述符子集搜索用于抗菌肽预测

Optimal Descriptor Subset Search via Chemical Information and Target Activity-Guided Algorithm for Antimicrobial Peptide Prediction.

作者信息

García-González Luis A, Marrero-Ponce Yovani, García-Jacas César R, Aguila Puentes Sergio A

机构信息

Centro de Nanociencias y Nanotecnología, Universidad Nacional Autónoma de México, Km. 107 Carretera Tijuana-Ensenada, Ensenada, Baja California C. P. 22860, México.

Facultad de Ingeniería. Universidad Panamericana. Augusto Rodin No. 498, Insurgentes Mixcoac, Benito Juárez, Ciudad de México 03920, México.

出版信息

J Chem Inf Model. 2025 Jul 14;65(13):6621-6631. doi: 10.1021/acs.jcim.5c00600. Epub 2025 Jun 18.

DOI:10.1021/acs.jcim.5c00600

PMID:40528473

Abstract

Antimicrobial peptides (AMPs) have emerged as a promising alternative to conventional drugs due to their potential applications in combating multidrug-resistant pathogens. Various computational approaches have been developed for AMP prediction, ranging from shallow learning methods to advanced deep learning techniques. Additionally, the performance of shallow learning models based on self-learning features derived from protein language models has recently been studied. However, the performance of AMP models based on shallow learning strongly depends on the quality of descriptors derived via manual feature engineering, which may miss crucial information by assuming that the initial descriptor set fully captures relevant information. The AExOp-DCS algorithm was introduced as an automatic feature domain optimization method that identifies the "optimal" descriptor set driven by the chemical structure and biological activity of the compounds under study. QSAR models built on AExOp-DCS optimized descriptors outperform those using nonoptimized sets. In this study, we explore the use of AExOp-DCS to identify optimal descriptor subsets for AMP modeling. Experimental results show that the descriptors returned by AExOp-DCS contain information comparable to those used in top-performing models while exhibiting higher discriminative capacity. The generated models based on the descriptors returned by AExOp-DCS achieved performance metric values comparable to state-of-the-art approaches while utilizing fewer descriptors, suggesting a more efficient modeling process. By reducing dimensionality without sacrificing accuracy, this approach contributes to the development of more efficient computational pipelines for AMP discovery. Finally, a Java software called AExOp-DCS-SEQ is freely available, enabling researchers to leverage its capabilities for peptide descriptor search and AMP classification tasks.

摘要

抗菌肽（AMPs）因其在对抗多重耐药病原体方面的潜在应用，已成为传统药物的一种有前景的替代品。已经开发了各种计算方法用于抗菌肽预测，从浅层学习方法到先进的深度学习技术。此外，最近还研究了基于从蛋白质语言模型衍生的自学习特征的浅层学习模型的性能。然而，基于浅层学习的抗菌肽模型的性能在很大程度上取决于通过手动特征工程获得的描述符的质量，而手动特征工程可能会因假设初始描述符集完全捕获了相关信息而遗漏关键信息。AExOp-DCS算法作为一种自动特征域优化方法被引入，该方法可识别由所研究化合物的化学结构和生物活性驱动的“最优”描述符集。基于AExOp-DCS优化描述符构建的定量构效关系（QSAR）模型优于使用未优化描述符集构建的模型。在本研究中，我们探索使用AExOp-DCS来识别用于抗菌肽建模的最优描述符子集。实验结果表明，AExOp-DCS返回的描述符所包含的信息与表现最佳的模型所使用的描述符相当，同时具有更高的区分能力。基于AExOp-DCS返回的描述符生成的模型在使用更少描述符的情况下，实现了与现有最先进方法相当的性能指标值，这表明建模过程更高效。通过在不牺牲准确性的情况下降低维度，这种方法有助于开发更高效的抗菌肽发现计算流程。最后，一款名为AExOp-DCS-SEQ的Java软件可免费获取，使研究人员能够利用其功能进行肽描述符搜索和抗菌肽分类任务。

相似文献

Optimal Descriptor Subset Search via Chemical Information and Target Activity-Guided Algorithm for Antimicrobial Peptide Prediction.

J Chem Inf Model. 2025 Jul 14;65(13):6621-6631. doi: 10.1021/acs.jcim.5c00600. Epub 2025 Jun 18.

AI-Driven Antimicrobial Peptide Discovery: Mining and Generation.

Acc Chem Res. 2025 Jun 17;58(12):1831-1846. doi: 10.1021/acs.accounts.0c00594. Epub 2025 Jun 3.

Short-Term Memory Impairment

Management of urinary stones by experts in stone disease (ESD 2025).

Arch Ital Urol Androl. 2025 Jun 30;97(2):14085. doi: 10.4081/aiua.2025.14085.

Comparison of Two Modern Survival Prediction Tools, SORG-MLA and METSSS, in Patients With Symptomatic Long-bone Metastases Who Underwent Local Treatment With Surgery Followed by Radiotherapy and With Radiotherapy Alone.

Clin Orthop Relat Res. 2024 Dec 1;482(12):2193-2208. doi: 10.1097/CORR.0000000000003185. Epub 2024 Jul 23.

Artificial intelligence for diagnosing exudative age-related macular degeneration.

Cochrane Database Syst Rev. 2024 Oct 17;10(10):CD015522. doi: 10.1002/14651858.CD015522.pub2.

Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.

Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.

Approaches for predicting dairy cattle methane emissions: from traditional methods to machine learning.

J Anim Sci. 2024 Jan 3;102. doi: 10.1093/jas/skae219.

Signs and symptoms to determine if a patient presenting in primary care or hospital outpatient settings has COVID-19.

Cochrane Database Syst Rev. 2022 May 20;5(5):CD013665. doi: 10.1002/14651858.CD013665.pub3.

[Volume and health outcomes: evidence from systematic reviews and from evaluation of Italian hospital data].

Epidemiol Prev. 2013 Mar-Jun;37(2-3 Suppl 2):1-100.

本文引用的文献

Machine learning for antimicrobial peptide identification and design.

Nat Rev Bioeng. 2024 May;2(5):392-407. doi: 10.1038/s44222-024-00152-x. Epub 2024 Feb 26.

AI Methods for Antimicrobial Peptides: Progress and Challenges.

Microb Biotechnol. 2025 Jan;18(1):e70072. doi: 10.1111/1751-7915.70072.

MD-LAIs Software: Computing Whole-Sequence and Amino Acid-Level "Embeddings" for Peptides and Proteins.

J Chem Inf Model. 2024 Dec 9;64(23):8665-8672. doi: 10.1021/acs.jcim.3c01189. Epub 2024 Nov 18.

QSAR modelling of enzyme inhibition toxicity of ionic liquid based on chaotic spotted hyena optimization algorithm.

SAR QSAR Environ Res. 2024 Sep;35(9):757-770. doi: 10.1080/1062936X.2024.2404853. Epub 2024 Sep 30.

Peptide-based drug discovery through artificial intelligence: towards an autonomous design of therapeutic peptides.

Brief Bioinform. 2024 May 23;25(4). doi: 10.1093/bib/bbae275.

Antimicrobial Peptides towards Clinical Application-A Long History to Be Concluded.

Int J Mol Sci. 2024 Apr 29;25(9):4870. doi: 10.3390/ijms25094870.

The role and future prospects of artificial intelligence algorithms in peptide drug development.

Biomed Pharmacother. 2024 Jun;175:116709. doi: 10.1016/j.biopha.2024.116709. Epub 2024 May 6.

Examining evolutionary scale modeling-derived different-dimensional embeddings in the antimicrobial peptide classification through a KNIME workflow.

Protein Sci. 2024 Apr;33(4):e4928. doi: 10.1002/pro.4928.

Artificial intelligence-driven antimicrobial peptide discovery.

Curr Opin Struct Biol. 2023 Dec;83:102733. doi: 10.1016/j.sbi.2023.102733. Epub 2023 Nov 21.

StarPep Toolbox: an open-source software to assist chemical space analysis of bioactive peptides and their functions using complex networks.

Bioinformatics. 2023 Aug 1;39(8). doi: 10.1093/bioinformatics/btad506.

文献AI研究员

20分钟写一篇综述，助力文献阅读效率提升50倍。

立即体验

用中文搜PubMed

大模型驱动的PubMed中文搜索引擎

马上搜索

文档翻译

学术文献翻译模型，支持多种主流文档格式。

立即体验

基于化学信息和靶标活性导向算法的最优描述符子集搜索用于抗菌肽预测

Optimal Descriptor Subset Search via Chemical Information and Target Activity-Guided Algorithm for Antimicrobial Peptide Prediction.

作者信息

机构信息

出版信息

相似文献

本文引用的文献

文献AI研究员

用中文搜PubMed

文档翻译

Suppr 超能文献

相似文献

本文引用的文献