Mengke Guo, Xiucai Ye, Dong Huang, Tetsuya Sakurai
{"title":"Robust feature learning using contractive autoencoders for multi-omics clustering in cancer subtyping.","authors":"Mengke Guo, Xiucai Ye, Dong Huang, Tetsuya Sakurai","doi":"10.1016/j.ymeth.2024.11.013","DOIUrl":null,"url":null,"abstract":"<p><p>Cancer can manifest in virtually any tissue or organ, necessitating precise subtyping of cancer patients to enhance diagnosis, treatment, and prognosis. With the accumulation of vast amounts of omics data, numerous studies have focused on integrating multi-omics data for cancer subtyping using clustering techniques. However, due to the heterogeneity of different omics data, extracting important features to effectively integrate these data for accurate clustering analysis remains a significant challenge. This study proposes a new multi-omics clustering framework for cancer subtyping, which utilizes contractive autoencoder to extract robust features. By encouraging the learned representation to be less sensitive to small changes, the contractive autoencoder learns robust feature representations from different omics. To incorporate survival information into the clustering analysis, Cox proportional hazards regression is used to further select the key features significantly associated with survival for integration. Finally, we utilize K-means clustering on the integrated feature to obtain the clustering result. The proposed framework is evaluated on ten different cancer datasets across four levels of omics data and compared to other existing methods. The experimental results indicate that the proposed framework effectively integrates the four omics datasets and outperforms other methods, achieving higher C-index scores and showing more significant differences between survival curves. Additionally, differential gene analysis and pathway enrichment analysis are performed to further demonstrate the effectiveness of the proposed method framework.</p>","PeriodicalId":390,"journal":{"name":"Methods","volume":" ","pages":""},"PeriodicalIF":4.2000,"publicationDate":"2024-11-20","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Methods","FirstCategoryId":"99","ListUrlMain":"https://doi.org/10.1016/j.ymeth.2024.11.013","RegionNum":3,"RegionCategory":"生物学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"BIOCHEMICAL RESEARCH METHODS","Score":null,"Total":0}
引用次数: 0
Abstract
Cancer can manifest in virtually any tissue or organ, necessitating precise subtyping of cancer patients to enhance diagnosis, treatment, and prognosis. With the accumulation of vast amounts of omics data, numerous studies have focused on integrating multi-omics data for cancer subtyping using clustering techniques. However, due to the heterogeneity of different omics data, extracting important features to effectively integrate these data for accurate clustering analysis remains a significant challenge. This study proposes a new multi-omics clustering framework for cancer subtyping, which utilizes contractive autoencoder to extract robust features. By encouraging the learned representation to be less sensitive to small changes, the contractive autoencoder learns robust feature representations from different omics. To incorporate survival information into the clustering analysis, Cox proportional hazards regression is used to further select the key features significantly associated with survival for integration. Finally, we utilize K-means clustering on the integrated feature to obtain the clustering result. The proposed framework is evaluated on ten different cancer datasets across four levels of omics data and compared to other existing methods. The experimental results indicate that the proposed framework effectively integrates the four omics datasets and outperforms other methods, achieving higher C-index scores and showing more significant differences between survival curves. Additionally, differential gene analysis and pathway enrichment analysis are performed to further demonstrate the effectiveness of the proposed method framework.
期刊介绍:
Methods focuses on rapidly developing techniques in the experimental biological and medical sciences.
Each topical issue, organized by a guest editor who is an expert in the area covered, consists solely of invited quality articles by specialist authors, many of them reviews. Issues are devoted to specific technical approaches with emphasis on clear detailed descriptions of protocols that allow them to be reproduced easily. The background information provided enables researchers to understand the principles underlying the methods; other helpful sections include comparisons of alternative methods giving the advantages and disadvantages of particular methods, guidance on avoiding potential pitfalls, and suggestions for troubleshooting.