|国家预印本平台
首页|Multi-omics Data Integration by Generative Adversarial Network

Multi-omics Data Integration by Generative Adversarial Network

Multi-omics Data Integration by Generative Adversarial Network

来源:bioRxiv_logobioRxiv
英文摘要

Accurate disease phenotype prediction plays an important role in the treatment of heterogeneous diseases like cancer in the era of precision medicine. With the advent of high throughput technologies, more comprehensive multi-omics data is now available that can effectively link the genotype to phenotype. However, the interactive relation of multi-omics datasets makes it particularly challenging to incorporate different biological layers to discover the coherent biological signatures and predict phenotypic outcomes. In this study, we introduce omicsGAN, a generative adversarial network (GAN) model to integrate two omics data and their interaction network. The model captures information from the interaction network as well as the two omics datasets and fuse them to generate synthetic data with better predictive signals. Large-scale experiments on The Cancer Genome Atlas (TCGA) breast cancer, lung cancer, and ovarian cancer datasets validate that (1) the model can effectively integrate two omics data (e.g., mRNA and microRNA expression data) and their interaction network (e.g., microRNA-mRNA interaction network). The synthetic omics data generated by the proposed model has a better performance on cancer outcome classification and patients survival prediction compared to original omics datasets. (2) The integrity of the interaction network plays a vital role in the generation of synthetic data with higher predictive quality. Using a random interaction network does not allow the framework to learn meaningful information from the omics datasets; therefore, results in synthetic data with weaker predictive signals.

Sun Jiao、Cheng Sze、Yong Jeongsik、Ahmed Khandakar Tanvir、Zhang Wei

Department of Computer Science, University of Central Florida||Genomics and Bioinformatics Cluster, University of Central FloridaDepartment of Biochemistry, Molecular Biology and Biophysics, University of Minnesota Twin CitiesDepartment of Biochemistry, Molecular Biology and Biophysics, University of Minnesota Twin CitiesDepartment of Computer Science, University of Central Florida||Genomics and Bioinformatics Cluster, University of Central FloridaDepartment of Computer Science, University of Central Florida||Genomics and Bioinformatics Cluster, University of Central Florida

10.1101/2021.03.13.435251

肿瘤学生物科学研究方法、生物科学研究技术计算技术、计算机技术

multiomics datadata integrationgenerative adversarial networkmultiomics interaction network

Sun Jiao,Cheng Sze,Yong Jeongsik,Ahmed Khandakar Tanvir,Zhang Wei.Multi-omics Data Integration by Generative Adversarial Network[EB/OL].(2025-03-28)[2025-05-04].https://www.biorxiv.org/content/10.1101/2021.03.13.435251.点此复制

评论