• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

一种基于Copula熵的高效交互式高维遗传数据特征选择方法。

An efficient and interactive feature selection approach based on copula entropy for high-dimensional genetic data.

作者信息

Yan Xiaoran, Shang Shilong, Li Dongxi, Dang Yun

机构信息

College of Artificial Intelligence, Taiyuan University of Technology, Taiyuan, Shanxi, China.

College of Computer Science and Technology, Taiyuan University of Technology, Taiyuan, Shanxi, China.

出版信息

Sci Rep. 2025 Aug 17;15(1):30100. doi: 10.1038/s41598-025-15068-8.

DOI:10.1038/s41598-025-15068-8
PMID:40820100
Abstract

Feature selection (FS) is especially important for high-dimensional data. In this paper, we propose an efficient and interactive feature selection approach based on copula entropy (CEFS+). The method combines feature-feature mutual information with feature-label mutual information and uses a maximum correlation minimum redundancy strategy for greedy selection. The approach uses copula entropy as a measure of feature relevance that captures the full-order interaction gain between features. Moreover, we prove the divisibility of multivariate mutual information, and derive a novel feature criterion, and propose a feature selection approach based on copula entropy called CEFS. Meanwhile, to overcome the instability of the CEFS method on some datasets, we propose the improved method CEFS+ which based on the rank technique. Finally, we evaluate the effectiveness of CEFS and CEFS+ using three classifiers on five datasets. In 10 out of 15 scenarios, our approach obtains the highest classification accuracy, which is much higher than the other six commonly used FS methods. In particular, our approach performs better on high-dimensional genetic datasets.

摘要

特征选择(FS)对于高维数据尤为重要。在本文中,我们提出了一种基于copula熵的高效交互式特征选择方法(CEFS+)。该方法将特征-特征互信息与特征-标签互信息相结合,并采用最大相关最小冗余策略进行贪心选择。该方法使用copula熵作为特征相关性的度量,以捕获特征之间的全阶交互增益。此外,我们证明了多元互信息的可分性,推导了一种新的特征准则,并提出了一种基于copula熵的特征选择方法CEFS。同时,为了克服CEFS方法在某些数据集上的不稳定性,我们提出了基于排序技术的改进方法CEFS+。最后,我们使用三个分类器在五个数据集上评估了CEFS和CEFS+的有效性。在15个场景中的10个场景中,我们的方法获得了最高的分类准确率,远高于其他六种常用的FS方法。特别是,我们的方法在高维遗传数据集上表现更好。

相似文献

1
An efficient and interactive feature selection approach based on copula entropy for high-dimensional genetic data.一种基于Copula熵的高效交互式高维遗传数据特征选择方法。
Sci Rep. 2025 Aug 17;15(1):30100. doi: 10.1038/s41598-025-15068-8.
2
Prescription of Controlled Substances: Benefits and Risks管制药品的处方:益处与风险
3
Classification of finger movements through optimal EEG channel and feature selection.通过最优脑电图通道和特征选择对手指运动进行分类。
Front Hum Neurosci. 2025 Jul 16;19:1633910. doi: 10.3389/fnhum.2025.1633910. eCollection 2025.
4
Signs and symptoms to determine if a patient presenting in primary care or hospital outpatient settings has COVID-19.在基层医疗机构或医院门诊环境中,如果患者出现以下症状和体征,可判断其是否患有 COVID-19。
Cochrane Database Syst Rev. 2022 May 20;5(5):CD013665. doi: 10.1002/14651858.CD013665.pub3.
5
Leveraging a foundation model zoo for cell similarity search in oncological microscopy across devices.利用基础模型库进行跨设备肿瘤显微镜检查中的细胞相似性搜索。
Front Oncol. 2025 Jun 18;15:1480384. doi: 10.3389/fonc.2025.1480384. eCollection 2025.
6
Systemic pharmacological treatments for chronic plaque psoriasis: a network meta-analysis.系统性药理学治疗慢性斑块状银屑病:网络荟萃分析。
Cochrane Database Syst Rev. 2021 Apr 19;4(4):CD011535. doi: 10.1002/14651858.CD011535.pub4.
7
An effective hybrid feature selection using entropy weight method for automatic sleep staging.一种基于熵权法的有效混合特征选择用于自动睡眠分期。
Physiol Meas. 2023 Oct 31;44(10). doi: 10.1088/1361-6579/acff35.
8
ASAS-NANP symposium: mathematical modeling in animal nutrition: synthetic database generation for non-normal multivariate distributions: a rank-based method with application to ruminant methane emissions.美国动物科学学会-北美猪营养大会研讨会:动物营养中的数学建模:非正态多元分布的综合数据库生成:一种基于秩的方法及其在反刍动物甲烷排放中的应用
J Anim Sci. 2025 Jan 4;103. doi: 10.1093/jas/skaf136.
9
Systemic pharmacological treatments for chronic plaque psoriasis: a network meta-analysis.慢性斑块状银屑病的全身药理学治疗:一项网状Meta分析。
Cochrane Database Syst Rev. 2020 Jan 9;1(1):CD011535. doi: 10.1002/14651858.CD011535.pub3.
10
Computer and mobile technology interventions for self-management in chronic obstructive pulmonary disease.用于慢性阻塞性肺疾病自我管理的计算机和移动技术干预措施。
Cochrane Database Syst Rev. 2017 May 23;5(5):CD011425. doi: 10.1002/14651858.CD011425.pub2.

本文引用的文献

1
CRIA: An Interactive Gene Selection Algorithm for Cancers Prediction Based on Copy Number Variations.CRIA:一种基于拷贝数变异的用于癌症预测的交互式基因选择算法。
Front Plant Sci. 2022 Mar 21;13:839044. doi: 10.3389/fpls.2022.839044. eCollection 2022.
2
Long non‑coding RNA suppresses ovarian cancer progression by regulating Hippo/YAP signaling.长链非编码 RNA 通过调控 Hippo/YAP 信号通路抑制卵巢癌细胞的进展。
Int J Mol Med. 2021 Apr;47(4). doi: 10.3892/ijmm.2021.4877. Epub 2021 Feb 12.
3
Integration of lncRNA and mRNA profiles to reveal the protective effects of extract on the gastrointestinal tract of mice subjected to D‑galactose‑induced aging.
整合长链非编码 RNA 和信使 RNA 谱揭示提取物对 D-半乳糖诱导衰老小鼠胃肠道的保护作用。
Int J Mol Med. 2021 Mar;47(3). doi: 10.3892/ijmm.2020.4834. Epub 2021 Jan 15.
4
Zinc finger and BTB domain-containing 7C (ZBTB7C) expression as an independent prognostic factor for colorectal cancer and its relevant molecular mechanisms.含锌指和BTB结构域蛋白7C(ZBTB7C)表达作为结直肠癌的独立预后因素及其相关分子机制
Am J Transl Res. 2020 Aug 15;12(8):4141-4159. eCollection 2020.
5
MiR-645 promotes invasiveness, metastasis and tumor growth in colorectal cancer by targeting EFNA5.miR-645 通过靶向 EFNA5 促进结直肠癌的侵袭、转移和肿瘤生长。
Biomed Pharmacother. 2020 May;125:109889. doi: 10.1016/j.biopha.2020.109889. Epub 2020 Feb 6.
6
MicroRNA-645 represses hepatocellular carcinoma progression by inhibiting SOX30-mediated p53 transcriptional activation.微小 RNA-645 通过抑制 SOX30 介导的 p53 转录激活抑制肝癌进展。
Int J Biol Macromol. 2019 Jan;121:214-222. doi: 10.1016/j.ijbiomac.2018.10.032. Epub 2018 Oct 9.
7
Multi-omic and multi-view clustering algorithms: review and cancer benchmark.多组学和多视角聚类算法:综述和癌症基准测试。
Nucleic Acids Res. 2018 Nov 16;46(20):10546-10562. doi: 10.1093/nar/gky889.
8
Downregulation of microRNA-645 suppresses breast cancer cell metastasis via targeting DCDC2.下调 microRNA-645 通过靶向 DCDC2 抑制乳腺癌细胞转移。
Eur Rev Med Pharmacol Sci. 2017 Sep;21(18):4129-4136.
9
Kr-POK (ZBTB7c) regulates cancer cell proliferation through glutamine metabolism.KR-POK(ZBTB7c)通过谷氨酰胺代谢调节癌细胞增殖。
Biochim Biophys Acta Gene Regul Mech. 2017 Aug;1860(8):829-838. doi: 10.1016/j.bbagrm.2017.05.005. Epub 2017 May 30.
10
Dysregulated miR-645 affects the proliferation and invasion of head and neck cancer cell.失调的miR-645影响头颈癌细胞的增殖和侵袭。
Cancer Cell Int. 2015 Sep 17;15:87. doi: 10.1186/s12935-015-0238-5. eCollection 2015.