Simultaneous Representation Learning of Multi-Omics and Clinical Outcome Data via a Supervised Knowledge-Guided Bayesian Factor Model

基于监督知识引导贝叶斯因子模型的多组学和临床结果数据同步表征学习

阅读:1

Abstract

With the advent of high-throughput techniques, multi-omics data and various clinical outcomes have been collected for a range of diseases. Multi-omics data play a crucial role in uncovering complex biological processes, yet simultaneous representation learning of such high-dimensional, heterogeneous multi-modality data along with clinical outcomes remains limited. To address this gap, we propose a supervised knowledge-guided Bayesian factor model for integrative analysis of multi-omics and clinical outcome data. The proposed method simultaneously extracts an informative low-dimensional representation and predicts one or more clinical outcomes of interest. The two-level adaptive shrinkage in the novel hierarchical priors allows for the identification of both active modalities and features, resulting in a biologically meaningful structural identification of the high-dimensional data. Moreover, the method is robust to noisy edges in biological graphs that do not align with ground truth. Finally, the proposed method can handle different data types including both continuous and categorical data. Extensive simulation studies and real data analyses of Alzheimer's disease (AD) data demonstrate the advantages of the proposed approach over existing methods. Notably, our analysis of multi-omics and imaging phenotype data from ADNI provides meaningful insights into the underlying biological mechanisms of AD.

特别声明

1、本页面内容包含部分的内容是基于公开信息的合理引用;引用内容仅为补充信息,不代表本站立场。

2、若认为本页面引用内容涉及侵权,请及时与本站联系,我们将第一时间处理。

3、其他媒体/个人如需使用本页面原创内容,需注明“来源:[生知库]”并获得授权;使用引用内容的,需自行联系原作者获得许可。

4、投稿及合作请联系:info@biocloudy.com。