Meng Zhang, Xiang Li, Adelumola Oladeinde, Michael Rothrock, Anthony Pokoo-Aikins, Gregory Zock
{"title":"A Novel Slope-Matrix-Graph Algorithm to Analyze Compositional Microbiome Data","authors":"Meng Zhang, Xiang Li, Adelumola Oladeinde, Michael Rothrock, Anthony Pokoo-Aikins, Gregory Zock","doi":"10.3390/microorganisms12091866","DOIUrl":null,"url":null,"abstract":"Networks are widely used to represent relationships between objects, including microorganisms within ecosystems, based on high-throughput sequencing data. However, challenges arise with appropriate statistical algorithms, handling of rare taxa, excess zeros in compositional data, and interpretation. This work introduces a novel Slope-Matrix-Graph (SMG) algorithm to identify microbiome correlations primarily based on slope-based distance calculations. SMG effectively handles any proportion of zeros in compositional data and involves: (1) searching for correlated relationships (e.g., positive and negative directions of changes) based on a “target of interest” within a setting, and (2) quantifying graph changes via slope-based distances between objects. Evaluations on simulated datasets demonstrated SMG’s ability to accurately cluster microbes into distinct positive/negative correlation groups, outperforming methods like Bray–Curtis and SparCC in both sensitivity and specificity. Moreover, SMG demonstrated superior accuracy in detecting differential abundance (DA) compared to ZicoSeq and ANCOM-BC2, making it a robust tool for microbiome analysis. A key advantage is SMG’s natural capacity to analyze zero-inflated compositional data without transformations. Overall, this simple yet powerful algorithm holds promise for diverse microbiome analysis applications.","PeriodicalId":18667,"journal":{"name":"Microorganisms","volume":null,"pages":null},"PeriodicalIF":4.1000,"publicationDate":"2024-09-09","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Microorganisms","FirstCategoryId":"99","ListUrlMain":"https://doi.org/10.3390/microorganisms12091866","RegionNum":2,"RegionCategory":"生物学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q2","JCRName":"MICROBIOLOGY","Score":null,"Total":0}
引用次数: 0
Abstract
Networks are widely used to represent relationships between objects, including microorganisms within ecosystems, based on high-throughput sequencing data. However, challenges arise with appropriate statistical algorithms, handling of rare taxa, excess zeros in compositional data, and interpretation. This work introduces a novel Slope-Matrix-Graph (SMG) algorithm to identify microbiome correlations primarily based on slope-based distance calculations. SMG effectively handles any proportion of zeros in compositional data and involves: (1) searching for correlated relationships (e.g., positive and negative directions of changes) based on a “target of interest” within a setting, and (2) quantifying graph changes via slope-based distances between objects. Evaluations on simulated datasets demonstrated SMG’s ability to accurately cluster microbes into distinct positive/negative correlation groups, outperforming methods like Bray–Curtis and SparCC in both sensitivity and specificity. Moreover, SMG demonstrated superior accuracy in detecting differential abundance (DA) compared to ZicoSeq and ANCOM-BC2, making it a robust tool for microbiome analysis. A key advantage is SMG’s natural capacity to analyze zero-inflated compositional data without transformations. Overall, this simple yet powerful algorithm holds promise for diverse microbiome analysis applications.
期刊介绍:
Microorganisms (ISSN 2076-2607) is an international, peer-reviewed open access journal which provides an advanced forum for studies related to prokaryotic and eukaryotic microorganisms, viruses and prions. It publishes reviews, research papers and communications. Our aim is to encourage scientists to publish their experimental and theoretical results in as much detail as possible. There is no restriction on the length of the papers. The full experimental details must be provided so that the results can be reproduced. Electronic files and software regarding the full details of the calculation or experimental procedure, if unable to be published in a normal way, can be deposited as supplementary electronic material.