PyNetCor：一个用于大规模相关性分析的高性能Python软件包。

PyNetCor: a high-performance Python package for large-scale correlation analysis.

作者信息

Long Shibin, Xia Yan, Liang Lifeng, Yang Ying, Xie Hailiang, Wang Xiaokai

机构信息

Department of Data Science, 01Life Institute, Shenzhen 518000, China.

State Key Laboratory of Pharmaceutical Biotechnology, Chemistry and Biomedicine Innovation Center (ChemBIC), School of Life Sciences, Nanjing University, Nanjing 210023, China.

出版信息

NAR Genom Bioinform. 2024 Dec 18;6(4):lqae177. doi: 10.1093/nargab/lqae177. eCollection 2024 Dec.

DOI:10.1093/nargab/lqae177

PMID:39703431

原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC11655297/

Abstract

The development of multi-omics technologies has generated an abundance of biological datasets, providing valuable resources for investigating potential relationships within complex biological systems. However, most correlation analysis tools face computational challenges when dealing with these high-dimensional datasets containing millions of features. Here, we introduce pyNetCor, a fast and scalable tool for constructing correlation networks on large-scale and high-dimensional data. PyNetCor features optimized algorithms for both full correlation coefficient matrix computation and top-k correlation search, outperforming other tools in the field in terms of runtime and memory consumption. It utilizes a linear interpolation strategy to rapidly estimate values and achieve false discovery rate control, demonstrating a speedup of over 110 times compared to existing methods. Overall, pyNetCor supports large-scale correlation analysis, a crucial foundational step for various bioinformatics workflows, and can be easily integrated into downstream applications to accelerate the process of extracting biological insights from data.

摘要

多组学技术的发展产生了大量的生物学数据集，为研究复杂生物系统中的潜在关系提供了宝贵资源。然而，大多数相关性分析工具在处理这些包含数百万个特征的高维数据集时面临计算挑战。在此，我们介绍pyNetCor，这是一种用于在大规模和高维数据上构建相关网络的快速且可扩展的工具。PyNetCor具有针对全相关系数矩阵计算和前k个相关性搜索的优化算法，在运行时间和内存消耗方面优于该领域的其他工具。它采用线性插值策略来快速估计值并实现错误发现率控制，与现有方法相比，速度提高了110倍以上。总体而言，pyNetCor支持大规模相关性分析，这是各种生物信息学工作流程的关键基础步骤，并且可以轻松集成到下游应用程序中，以加速从数据中提取生物学见解的过程。

https://cdn.ncbi.nlm.nih.gov/pmc/blobs/685b/11655297/1e76e63b21e8/lqae177figgra1.jpg

相似文献

PyNetCor: a high-performance Python package for large-scale correlation analysis.PyNetCor：一个用于大规模相关性分析的高性能Python软件包。

NAR Genom Bioinform. 2024 Dec 18;6(4):lqae177. doi: 10.1093/nargab/lqae177. eCollection 2024 Dec.

A fast, scalable and versatile tool for analysis of single-cell omics data.一种快速、可扩展且功能多样的单细胞组学数据分析工具。

Nat Methods. 2024 Feb;21(2):217-227. doi: 10.1038/s41592-023-02139-9. Epub 2024 Jan 8.

Large-scale correlation network construction for unraveling the coordination of complex biological systems.大规模相关网络构建，用于揭示复杂生物系统的协调关系。

Nat Comput Sci. 2023 Apr;3(4):346-359. doi: 10.1038/s43588-023-00429-y. Epub 2023 Apr 13.

AnnSQL: A Python SQL-based package for fast large-scale single-cell genomics analysis using minimal computational resources.AnnSQL：一个基于Python和SQL的软件包，用于使用最少的计算资源进行快速大规模单细胞基因组学分析。

bioRxiv. 2025 Mar 22:2024.11.02.621676. doi: 10.1101/2024.11.02.621676.

SnapATAC2: a fast, scalable and versatile tool for analysis of single-cell omics data.SnapATAC2：一种用于单细胞组学数据分析的快速、可扩展且通用的工具。

bioRxiv. 2023 Sep 15:2023.09.11.557221. doi: 10.1101/2023.09.11.557221.

Ibaqpy: A scalable Python package for baseline quantification in proteomics leveraging SDRF metadata.Ibaqpy：一个用于蛋白质组学中利用SDRF元数据进行基线定量的可扩展Python软件包。

J Proteomics. 2025 Jun 15;317:105440. doi: 10.1016/j.jprot.2025.105440. Epub 2025 Apr 21.

A python library for the fast and scalable computation of biologically meaningful individual specific networks.一个用于快速和可扩展地计算具有生物学意义的个体特定网络的 Python 库。

Sci Rep. 2024 Aug 6;14(1):18243. doi: 10.1038/s41598-024-69067-2.

oFVSD: a Python package of optimized forward variable selection decoder for high-dimensional neuroimaging data.oFVSD：用于高维神经成像数据的优化前向变量选择解码器的Python软件包。

Front Neuroinform. 2023 Sep 26;17:1266713. doi: 10.3389/fninf.2023.1266713. eCollection 2023.

rstoolbox - a Python library for large-scale analysis of computational protein design data and structural bioinformatics.rstoolbox - 一个用于大规模分析计算蛋白质设计数据和结构生物信息学的 Python 库。

BMC Bioinformatics. 2019 May 15;20(1):240. doi: 10.1186/s12859-019-2796-3.

DNEA: an R package for fast and versatile data-driven network analysis of metabolomics data.DNEA：一个用于代谢组学数据快速且通用的数据驱动网络分析的R软件包。

BMC Bioinformatics. 2024 Dec 18;25(1):383. doi: 10.1186/s12859-024-05994-1.

本文引用的文献

Large-scale correlation network construction for unraveling the coordination of complex biological systems.大规模相关网络构建，用于揭示复杂生物系统的协调关系。

Nat Comput Sci. 2023 Apr;3(4):346-359. doi: 10.1038/s43588-023-00429-y. Epub 2023 Apr 13.

Gut Microbial Metabolite Butyrate and Its Therapeutic Role in Inflammatory Bowel Disease: A Literature Review.肠道微生物代谢物丁酸盐及其在炎症性肠病中的治疗作用：文献综述。

Nutrients. 2023 May 11;15(10):2275. doi: 10.3390/nu15102275.

Methods and applications for single-cell and spatial multi-omics.单细胞和空间多组学的方法和应用。

Nat Rev Genet. 2023 Aug;24(8):494-515. doi: 10.1038/s41576-023-00580-2. Epub 2023 Mar 2.

Dual proteome-scale networks reveal cell-specific remodeling of the human interactome.双重蛋白质组尺度网络揭示了人类相互作用组的细胞特异性重塑。

Cell. 2021 May 27;184(11):3022-3040.e28. doi: 10.1016/j.cell.2021.04.011. Epub 2021 May 6.

Integrating taxonomic, functional, and strain-level profiling of diverse microbial communities with bioBakery 3.利用 bioBakery 3 整合具有分类学、功能和菌株水平特征的多样化微生物群落。

Elife. 2021 May 4;10:e65088. doi: 10.7554/eLife.65088.

Butyrate mediates anti-inflammatory effects of in intestinal epithelial cells through .丁酸盐通过介导在肠道上皮细胞中的抗炎作用。

Gut Microbes. 2020 Nov 9;12(1):1-16. doi: 10.1080/19490976.2020.1826748.

Array programming with NumPy.使用 NumPy 进行数组编程。

Nature. 2020 Sep;585(7825):357-362. doi: 10.1038/s41586-020-2649-2. Epub 2020 Sep 16.

Earth microbial co-occurrence network reveals interconnection pattern across microbiomes.地球微生物共同发生网络揭示了微生物组之间的相互关联模式。

Microbiome. 2020 Jun 4;8(1):82. doi: 10.1186/s40168-020-00857-2.

Therapeutic Opportunities in Inflammatory Bowel Disease: Mechanistic Dissection of Host-Microbiome Relationships.炎症性肠病的治疗机会：宿主-微生物组关系的机制剖析。

Cell. 2019 Aug 22;178(5):1041-1056. doi: 10.1016/j.cell.2019.07.045.

Use and abuse of correlation analyses in microbial ecology.微生物生态学中相关性分析的使用与滥用。

ISME J. 2019 Nov;13(11):2647-2655. doi: 10.1038/s41396-019-0459-z. Epub 2019 Jun 28.

文献检索

告别复杂PubMed语法，用中文像聊天一样搜索，搜遍4000万医学文献。AI智能推荐，让科研检索更轻松。

立即免费搜索

文件翻译

保留排版，准确专业，支持PDF/Word/PPT等文件格式，支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述，25分钟生成高质量综述，智能提取关键信息，辅助科研写作。

立即免费体验

PyNetCor：一个用于大规模相关性分析的高性能Python软件包。

PyNetCor: a high-performance Python package for large-scale correlation analysis.

作者信息

机构信息

出版信息

相似文献

本文引用的文献

文献检索

文件翻译

深度研究

Suppr 超能文献

相似文献

本文引用的文献