Unverified paper record
Trait-based machine learning modeling of soluble carbohydrate content in Trachyspermum ammi exposed to organic and chemical fertilization and salicylic acid.
BMC plant biology · 9 Jul 2026 · 10.1186/s12870-026-09466-x
Abstract
Accurate assessment of biochemical traits in medicinal plants is essential for supporting environmentally responsible agriculture, improving crop quality, and enhancing the nutritional and pharmacological value of plant-derived products. Although medicinal plants are rich in bioactive compounds, conventional methods for measuring key biochemical components, such as soluble carbohydrates, are often time-consuming, destructive, and resource-intensive. Trachyspermum ammi L. (Ajwain) is valued for its antioxidant, antimicrobial, and digestive properties, highlighting the need for rapid, reliable, and non-destructive evaluation methods. Despite previous studies on fertilization effects on growth and bioactive compounds in T. ammi, research integrating morpho-physiological data with machine learning to predict key biochemical traits remains limited. In this study, we applied Multilayer Perceptron (MLP) and Gaussian Process Regression (GPR) models to estimate soluble carbohydrate content in a non-invasive and efficient manner. A dataset including morphological, biochemical, physiological, and macronutrient traits was used as input variables. Fertilization regimes and salicylic acid (SA) treatments were applied to induce variability in plant traits but were not directly included as model features, ensuring that predictions were trait-based. Models were trained and evaluated on n = 45 samples using five-fold cross-validation. Among the tested models, MLP and GPR achieved the highest predictive accuracy, particularly when the full feature set was used. Predictions based solely on biochemical and physiological traits were nearly as accurate as those using all variables, suggesting that these traits provide reliable and cost-effective estimates. Considering the limited dataset, results should be interpreted with caution, and future studies using larger, independent datasets are recommended to further assess model robustness and generalizability. These findings demonstrate the practical potential of the proposed machine learning approach for rapid, non-destructive assessment of biochemical traits in medicinal plants and may inform the development of GUI-based decision-support tools for precision agriculture and phytopharmaceutical research.
Plant phenotyping relevance
機械学習モデルによる植物の可溶性炭水化物含量の非破壊推定が研究の中心であり、交差検証による技術評価も行っているため、植物フェノタイピング手法として採用する。
abstractIn this study, we applied Multilayer Perceptron (MLP) and Gaussian Process Regression (GPR) models to estimate soluble carbohydrate content in a non-invasive and efficient manner.
abstractModels were trained and evaluated on n = 45 samples using five-fold cross-validation.
abstractThese findings demonstrate the practical potential of the proposed machine learning approach for rapid, non-destructive assessment of biochemical traits in medicinal plants
Code and data availability
The supplied blocks contain no data availability, code availability, or repository statements. The phenotype dataset (n=45 trait measurements) and MATLAB R2019b model implementations are described but no public deposit, URL, or availability language appears anywhere in the text. The only URL present (https://doi.org/10
No evidence-backed public reproduction asset is currently recorded.
This is an automatically classified, unverified record. Curator approval is required before any resource enters the Catalog.