Predicting Fitness-Related Traits Using Gene Expression and Machine Learning

利用基因表达和机器学习预测与适应性相关的特征

阅读:1

Abstract

Evolution by natural selection occurs at its most basic through the change in frequencies of alleles; connecting those genomic targets to phenotypic selection is an important goal for evolutionary biology in the genomics era. The relative abundance of gene products expressed in a tissue can be considered a phenotype intermediate to the genes and genomic regulatory elements themselves and more traditionally measured macroscopic phenotypic traits such as flowering time, size, or growth. The high dimensionality, low sample size nature of transcriptomic sequence data is a double-edged sword, however, as it provides abundant information but makes traditional statistics difficult. Machine learning (ML) has many features which handle high-dimensional data well and is thus useful in genetic sequence applications. Here, we examined the association of fitness components with gene expression data in Ipomoea hederacea (Ivyleaf morning glory) grown under field conditions. We combine the results of two different ML approaches and find evidence that expression of photosynthesis-related genes is likely under selection. We also find that genes related to stress and light responses were overall important in predicting fitness. With this study, we demonstrate the utility of ML models for smaller samples and their potential application for understanding natural selection.

特别声明

1、本页面内容包含部分的内容是基于公开信息的合理引用;引用内容仅为补充信息,不代表本站立场。

2、若认为本页面引用内容涉及侵权,请及时与本站联系,我们将第一时间处理。

3、其他媒体/个人如需使用本页面原创内容,需注明“来源:[生知库]”并获得授权;使用引用内容的,需自行联系原作者获得许可。

4、投稿及合作请联系:info@biocloudy.com。