LPItabformer: Enhancing generalization in predicting lncRNA-protein interactions via a tabular Transformer

LPItabformer:通过表格化Transformer增强预测lncRNA-蛋白质相互作用的泛化能力

阅读:3

Abstract

Long-noncoding RNAs (LncRNAs) play important roles in physiological and pathological processes. Accurately predicting lncRNA-protein interactions (LPIs) is vital strategy for clarify functions and pathogenic mechanisms of lncRNAs. Current computational methods for evaluating LPIs with their utility and generalization have significant room for improvement. In this study, data splitting by incorporating protein clusters as group information reveals that lots of LPI prediction methods suffer from generalization flaws due to data leakage caused by ignoring LPI biological properties. To address the issue, we present LPItabformer, a tabular Transformer framework for predicting LPIs, that incorporates a domain shifts with uncertainty (DSU) module for generalization enhancement. The LPItabformer demonstrates a capacity to alleviate the generalization challenges associated with biases in LPI data and preferences in protein binding patterns. In addition, LPItabformer shows greater robustness and generalization on human and mouse LPI datasets compared to state-of-the-art methods. Ultimately, we have verified that the LPItabformer is capable of predicting novel LPIs. Code is available at https://github.com/Ci-TJ/LPItabformer.

特别声明

1、本页面内容包含部分的内容是基于公开信息的合理引用;引用内容仅为补充信息,不代表本站立场。

2、若认为本页面引用内容涉及侵权,请及时与本站联系,我们将第一时间处理。

3、其他媒体/个人如需使用本页面原创内容,需注明“来源:[生知库]”并获得授权;使用引用内容的,需自行联系原作者获得许可。

4、投稿及合作请联系:info@biocloudy.com。