All source codes are available from the repository in GitHub: https://github.com/Yoska393/Twostep .
Open resource ↗Yoska393/Twostep · lines:306-317Unverified paper record
Integration of proxy intermediate omics traits into a nonlinear two-step model for accurate phenotypic prediction.
TAG. Theoretical and applied genetics. Theoretische und angewandte Genetik · 5 Mar 2026 · 10.1007/s00122-026-05171-3
Abstract
Intermediate omics traits, which mediate the effects of genetic variation on phenotypic traits, are increasingly recognized as valuable components of genetic evaluation. In particular, rhizosphere microbiota play a crucial role in plant health and productivity; however, their complex interactions with host genetics remain challenging to model. Although two-step modeling frameworks have been proposed to integrate intermediate omics traits into phenotype prediction, existing approaches do not incorporate nonlinear relationships between different omics layers. To address this, we have proposed a two-step phenotype prediction framework that integrates genomic, rhizosphere microbiome, and metabolome (meta-metabolome) data, while explicitly capturing omics-omics nonlinearities. The first step is to predict meta-metabolome traits from genetic and microbial features, thus effectively isolating them from the environmental noise. In this process, intermediate "proxy" omics traits are generated as general biological information to provide robust models. The second step utilizes this "proxy" to enhance the accuracy of the phenotype prediction. We compared a linear mixed model (Best Linear Unbiased Prediction, BLUP) and a nonlinear model (Random Forest, RF) at each step, as demonstrated through simulations and empirical analysis of a multi-omics soybean dataset in which nonlinear modeling captures intricate omics interactions. Notably, our approach enables phenotype prediction without requiring the original meta-metabolome data used in model training, thereby reducing reliance on costly omics measurements. This framework integrates intermediate omics traits into genomic prediction to improve prediction accuracy and provide solutions for deeper insights into plant-microbiome interactions.
Plant phenotyping relevance
植物の表現型予測を目的とする非線形マルチオミクス計算フレームワークが研究の中心であり、単なるオミクス測定や生物学的実験ではない。
abstractwe have proposed a two-step phenotype prediction framework that integrates genomic, rhizosphere microbiome, and metabolome (meta-metabolome) data, while explicitly capturing omics-omics nonlinearities.
abstractThe second step utilizes this "proxy" to enhance the accuracy of the phenotype prediction.
Code and data availability
The paper's analysis code is publicly available on GitHub, and the metabolome data are publicly available via the RIKEN DropMet website (IDs DM0071, DM0072). Phenotype and other multi-omics data are only available from the corresponding author upon request.
This is an automatically classified, unverified record. Curator approval is required before any resource enters the Catalog.