Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and management response. Modern 3D reconstruction methods perform strongly on generic benchmarks, but rendered appearance may not translate into metrically and agronomically useful geometry in crop fields. We introduce UAV3DCrop, a public benchmark of repeated multi-angle unmanned aerial vehicle (UAV) crop surveys. It contains 88,830 RGB images at $5280 \times 3956$ pixels, with a ground sampling distance of 3.6-5.8 mm, from 91 scenes spanning corn, soybean, wheat, and oat. Track A evaluates seven scene-optimized methods -- Neural Radiance Field (NeRF) and 3D Gaussian Splatting (3DGS) variants -- on held-out views, photogrammetry-referenced depth, and canopy-height recovery. Track B tests four pretrained feed-forward models on zero-shot camera-pose and geometry estimation. The scene-optimized methods rank differently across the three targets: Splatfacto-big leads appearance, whereas Scaffold-GS leads depth and is statistically tied with Splatfacto for canopy height. Among feed-forward models, MapAnything leads on seven of the eight metrics, while the remaining models vary more across crops and fail severely on absolute scale in a way that alignment conceals. Repeated acquisitions reveal further sensitivities that differ by output type and by model, associated with position within the acquisition sequence and with tie-point multiplicity. Current 3D reconstruction methods are therefore not yet interchangeable for agronomic use: no single method wins on appearance, geometry, and canopy height at once, and only one of four feed-forward models recovers usable metric scale. The dataset is publicly available at https://link-dev.github.io/UAV3DCrop/
Why it matches plant phenotyping methods植物キャノピー高さという明示的な形質を対象に、UAV 3D再構成手法をベンチマークし、公開データセットとして提供しているため、フェノタイピング手法が中心である。
abstractWe introduce UAV3DCrop, a public benchmark of repeated multi-angle unmanned aerial vehicle (UAV) crop surveys.
Reproduction assets foundThe paper introduces UAV3DCrop, a public benchmark of repeated multi-angle UAV crop surveys (88,830 RGB images, 91 scenes, four crops) with refined poses, photogrammetric depth references, and linked canopy-height and effective-LAI field measurements. The dataset is explicitly stated to be publicly available under CC BDataset · publiche acquisition sequence and with tie-point multiplicity. Current 3D reconstruction methods are therefore not yet interchangeable for agronomic use: no single method wins on appearance, geometry, and canopy height at once, and only one of four feed-forward models recovers usable metric scale. The dataset is publicly available at https://link-dev.github.io/UAV3DCrop/ .
Keywords:
UAV imagery; agricultural datasets; crop-field reconstruction; neural radiance fields;
Gaussian splatting; feed-forward geometry.
1 IntroductionOpen asset ↗UAV3DCroplines:1-90Code / dataset availability confirmedarXiv · OpenAlex · checked 11 Sept 2026
Tomographic microscopy enables three-dimensional internal imaging but often requires expensive optical or X-ray instrumentation. Here we present an ultra-low-cost continuous-wave diffusive tomography (CWDT) system for biological samples. The system uses a smartphone microscope, a white LED coupled into an optical fiber, 3D-printed micropositioners, and a physics-based forward model optimized with machine learning. We demonstrate full-color volumetric reconstructions from a tartrazine-cleared poplar section, a scattering phantom, fungal mycelium near an Arabidopsis root, and thick poplar branch imaging with an inserted side-emitting fiber. The current results are qualitative and exploratory, but they show that scanned fiber illumination and inexpensive hardware can produce useful three-dimensional reconstruction outputs for low-cost microscopy experiments.
Why it matches plant phenotyping methods低コスト三次元断層イメージング法そのものを開発し、ポプラ組織・枝やシロイヌナズナ根近傍を対象に植物の内部構造を可視化しているため、植物形態の取得法として中心的です。
abstractHere we present an ultra-low-cost continuous-wave diffusive tomography (CWDT) system for biological samples.
Reproduction assets foundThe paper's raw imaging inputs, configurations, and reconstruction outputs for Figures 2–5 are publicly deposited on Kaggle. The analysis code repository is only 'prepared for release' (no confirmed public deposit yet), so it is listed as request-only. Hardware CAD mirrors are public but are instrument designs, not theDataset · publicFigure-level raw inputs, model configurations, selected outputs, and manifests are available through the Kaggle dataset https://www.kaggle.com/datasets/alingold/continuous-wave-diffusive-tomography .Open asset ↗continuous-wave-diffusive-tomographylines:108-129Code / dataset availability confirmedarXiv · checked 13 Sept 2026
Natural and anthropogenic disturbances are impacting the health of forests worldwide. Monitoring forest disturbances at scale is important to inform conservation efforts. Here, we present a scalable approach for country-wide mapping of forest greenness anomalies at the 10 m resolution of Sentinel-2. Using relevant ecological and topographical context and an established representation of the vegetation cycle, we learn a predictive quantile model of the normalised difference vegetation index (NDVI) derived from Sentinel-2 data. The resulting expected seasonal cycles are used to detect NDVI anomalies across Switzerland between April 2017 and August 2025. Goodness-of-fit evaluations show that the conditional model explains 65% of the observed variations in the median seasonal cycle. The model consistently benefits from the local context information, particularly during the green-up period. The approach produces coherent spatial anomaly patterns and enables country-wide quantification of forest browning. Case studies with independent reference data from known events illustrate that the model reliably detects different types of disturbances.
Why it matches plant phenotyping methodsSentinel-2 NDVIを用いて森林キャノピーの季節変動から褐変・攪乱状態を推定する手法を開発し、適合度と独立参照データで検証しているため、単なる森林地図作成ではなく植物状態の取得・評価が中心である。
abstractwe present a scalable approach for country-wide mapping of forest greenness anomalies at the 10 m resolution of Sentinel-2.
Reproduction assets foundThe paper explicitly states that its code and interactive content are publicly available in the authors' GitHub repository. Other URLs in the article are cited third-party data sources (swisstopo, EnviDat, GDAL, TauDEM, WhiteboxTools) rather than paper-specific assets.Code · publicThe code and interactive content are available at https://github.com/SamanthaBiegel/s2-forest-browning-monitoring .Open asset ↗SamanthaBiegel/s2-forest-browning-monitoringlines:51-55Code / dataset availability confirmedOpenAlex · arXiv · checked 13 Sept 2026
Semantic reconstruction of agricultural scenes plays a vital role in tasks such as phenotyping and yield estimation. However, traditional approaches based on manual scanning or fixed camera setups remain a major bottleneck, while active-mapping methods based solely on occupancy grids are too coarse for accurate trait estimation. To address this gap, we propose an active 3D reconstruction framework for horticultural environments using a mobile manipulator. The system integrates OctoMap with 3D Gaussian Splatting to enable accurate and efficient target-aware mapping. A low-resolution OctoMap provides probabilistic occupancy information for informative viewpoint selection and collision-free planning, while 3D Gaussian Splatting leverages geometric, photometric, and semantic information to optimize 3D Gaussians for high-fidelity scene reconstruction. We further introduce a robust mapping strategy that mitigates semantic segmentation and depth noise, together with a background pruning method that reduces memory and computational cost. We validate our framework across simulated, laboratory, and real greenhouse scenes, showing consistent improvements across three state-of-the-art Gaussian Splatting backbones. In simulation, where ground-truth geometry is available, our approach outperforms occupancy-based mapping in both reconstruction accuracy and runtime efficiency: compared with a 0.01m-resolution OctoMap, it doubles the fruit-level F1 score under noisy conditions while achieving up to a threefold reduction in runtime. Beyond simulation, novel-view synthesis quality also improves consistently in laboratory and real greenhouse environments, with PSNR and mIoU improving by up to 1.5 dB and 18%, respectively. Finally, the reconstructed semantic maps enable fruit counting and volume estimation with accuracies approaching 80%.
Why it matches plant phenotyping methods園芸ロボット向けの3D再構成・能動マッピング手法を開発し、果実の計数・体積推定という植物形質の取得に適用・検証しているため、フェノタイピング手法が中心的です。
titleOctoSplat: Hybrid OctoMap-Gaussian Splatting for Active Semantic Mapping and Phenotyping with Horticultural Robots
Reproduction assets foundThe paper's supplementary material is hosted on the authors' public project page (jrcuaranv.github.io/octosplat), and the authors state that all code and data are publicly available. The SimSense repository is a third-party depth-sensor simulator tool, not a paper-specific asset.Code · publicAll code and data are publicly available to facilitate reproducibility.Open asset ↗lines:59-163Code / dataset availability confirmedarXiv · OpenAlex · checked 15 Sept 2026
AppleCottonPearField / plotNeRF / 3D Gaussian SplattingFruitCounting2D/3D reconstructionSegmentation
Rigorous crop counting is crucial for effective agricultural management and informed intervention strategies. However, in outdoor field environments, partial occlusions combined with inherent ambiguity in distinguishing clustered crops from individual viewpoints poses an immense challenge for image-based segmentation methods. To address these problems, we introduce a novel crop counting framework designed for exact enumeration via 3D instance segmentation. Our approach utilizes 2D images captured from multiple viewpoints and associates independent instance masks for neural radiance field (NeRF) view synthesis. We introduce crop visibility and mask consistency scores, which are incorporated alongside 3D information from a NeRF model. This results in an effective segmentation of crop instances in 3D and highly-accurate crop counts. Furthermore, our method eliminates the dependence on crop-specific parameter tuning. We validate our framework on three agricultural datasets consisting of cotton bolls, apples, and pears, and demonstrate consistent counting performance despite major variations in crop color, shape, and size. A comparative analysis against the state of the art highlights superior performance on crop counting tasks. Lastly, we contribute a cotton plant dataset to advance further research on this topic.
Why it matches plant phenotyping methodsNeRFと3Dインスタンスセグメンテーションを用いて作物個体・器官数を推定する画像ベース表現型計測手法を開発・検証しており、方法が研究の中心である。
abstractwe introduce a novel crop counting framework designed for exact enumeration via 3D instance segmentation.
Reproduction assets foundThe paper contributes a public infield cotton plant dataset (8 plants, ~150 iPhone images each, ground-truth boll counts, SAM instance masks) and states that source code, dataset, and multimedia are available at the authors' public project page, which is an allowed URL. The spectacularai GitHub URL is a generic third-pDataset · publicthat incorporates crop visibility
and mask consistency, enabling robustness against occlusions and annotation
discrepancies.
•
We release a public infield cotton plant dataset designed for 3D
rendering and cotton boll counting tasks.
The source code, dataset, and multimedia material associated with this project
can be found at
https://robotic-vision-lab.github.io/cropnerf .
II Related Work
II-A Image-Based Techniques
Image-based methods typically employ object detection to identify crops within
images. For example, Chen et al. [ 4 ] utilized multiple
convolutional neural networks (CNNs) to map input images to total fruit counts.
Similarly, Häni et al. [ 5 ] formulated crop counting as a
multOpen asset ↗lines:108-187Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
Potato yield is a key indicator for optimizing cultivation practices in agriculture. Potato yield can be estimated on harvesters using RGB-D cameras, which capture three-dimensional (3D) information of individual tubers moving along the conveyor belt. However, point clouds reconstructed from RGB-D images are incomplete due to self-occlusion, leading to systematic underestimation of tuber weight. To address this, we introduce PointRAFT, a high-throughput point cloud regression network that directly predicts continuous 3D shape properties, such as tuber weight, from partial point clouds. Rather than reconstructing full 3D geometry, PointRAFT infers target values directly from raw 3D data. Its key architectural novelty is an object height embedding that incorporates tuber height as an additional geometric cue, improving weight prediction under practical harvesting conditions. PointRAFT was trained and evaluated on 26,688 partial point clouds collected from 859 potato tubers across four cultivars and three growing seasons on an operational harvester in Japan. On a test set of 5,254 point clouds from 172 tubers, PointRAFT achieved a mean absolute error of 12.0 g and a root mean squared error of 17.2 g, substantially outperforming a linear regression baseline and a standard PointNet++ regression network. With an average inference time of 6.3 ms per point cloud, PointRAFT supports processing rates of up to 150 tubers per second, meeting the high-throughput requirements of commercial potato harvesters. Beyond potato weight estimation, PointRAFT provides a versatile regression network applicable to a wide range of 3D phenotyping and robotic perception tasks. The code, network weights, and a subset of the dataset are publicly available at https://github.com/pieterblok/pointraft.git.
Why it matches plant phenotyping methods部分点群からジャガイモ塊茎重量を推定する3D深層学習手法を開発・評価しており、植物形質取得が研究の中心である。
abstractwe introduce PointRAFT, a high-throughput point cloud regression network that directly predicts continuous 3D shape properties, such as tuber weight, from partial point clouds.
Reproduction assets foundThe paper publicly releases its authors' analysis code and trained network weights on GitHub, and a subset of its potato tuber partial point cloud dataset (with ground truth weights) on Hugging Face. Both are paper-specific, public, and actionable.Code · publicThe code, network weights, and a subset of the dataset are publicly available at https://github.com/pieterblok/pointraft.git .Open asset ↗pieterblok/pointraftlines:1-93Dataset · publicA subset of the datasets generated and/or analyzed during this study is publicly available at: https://huggingface.co/datasets/UTokyo-FieldPhenomics-Lab/3DPotatoTwinOpen asset ↗UTokyo-FieldPhenomics-Lab/3DPotatoTwinlines:447-463Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
Automating tasks in orchards is challenging because of the large amount of variation in the environment and occlusions. One of the challenges is apple pose estimation, where key points, such as the calyx, are often occluded. Recently developed pose estimation methods no longer rely on these key points, but still require them for annotations, making annotating challenging and time-consuming. Due to the abovementioned occlusions, there can be conflicting and missing annotations of the same fruit between different images. Novel 3D reconstruction methods can be used to simplify annotating and enlarge datasets. We propose a novel pipeline consisting of 3D Gaussian Splatting to reconstruct an orchard scene, simplified annotations, automated projection of the annotations to images, and the training and evaluation of a pose estimation method. Using our pipeline, 105 manual annotations were required to obtain 28,191 training labels, a reduction of 99.6%. Experimental results indicated that training with labels of fruits that are $\leq95\%$ occluded resulted in the best performance, with a neutral F1 score of 0.927 on the original images and 0.970 on the rendered images. Adjusting the size of the training dataset had small effects on the model performance in terms of F1 score and pose estimation accuracy. It was found that the least occluded fruits had the best position estimation, which worsened as the fruits became more occluded. It was also found that the tested pose estimation method was unable to correctly learn the orientation estimation of apples.
Why it matches plant phenotyping methodsリンゴの姿勢推定アノテーションを大幅に効率化する3D再構成・自動ラベル投影パイプラインを開発し、姿勢推定性能も評価しているため、植物フェノタイピング手法が中心である。
abstractWe propose a novel pipeline consisting of 3D Gaussian Splatting to reconstruct an orchard scene, simplified annotations, automated projection of the annotations to images, and the training and evaluation of a pose estimation method.
Reproduction assets foundThe paper explicitly provides two paper-specific public assets: the authors' phenotyping/pose-estimation pipeline code on GitHub and the collected apple orchard image dataset on a 4TU DOI. Both are directly used for the paper's measurements and analysis.Dataset · publicIn total, 367 images were collected. The dataset is available at https://doi.org/10.4121/976c94f2-028f-4291-adfd-20eb82b0f647Open asset ↗10.4121/976c94f2-028f-4291-adfd-20eb82b0f647lines:92-108Code / dataset availability confirmedarXiv · OpenAlex · checked 15 Sept 2026
Lychee is a high-value subtropical fruit. The adoption of vision-based harvesting robots can significantly improve productivity while reduce reliance on labor. High-quality data are essential for developing such harvesting robots. However, there are currently no consistently and comprehensively annotated open-source lychee datasets featuring fruits in natural growing environments. To address this, we constructed a dataset to facilitate lychee detection and maturity classification. Color (RGB) images were acquired under diverse weather conditions, and at different times of the day, across multiple lychee varieties, such as Nuomici, Feizixiao, Heiye, and Huaizhi. The dataset encompasses three different ripeness stages and contains 11,414 images, consisting of 878 raw RGB images, 8,780 augmented RGB images, and 1,756 depth images. The images are annotated with 9,658 pairs of lables for lychee detection and maturity classification. To improve annotation consistency, three individuals independently labeled the data, and their results were then aggregated and verified by a fourth reviewer. Detailed statistical analyses were done to examine the dataset. Finally, we performed experiments using three representative deep learning models to evaluate the dataset. It is publicly available for academic
Why it matches plant phenotyping methodsライチ果実の成熟段階という植物器官の状態をRGB-D画像から分類するデータセットを構築し、アノテーション検証と深層学習モデル評価を行っており、表現型取得・評価手法が中心である。
abstractwe constructed a dataset to facilitate lychee detection and maturity classification.
Reproduction assets foundThe authors publicly release the paper's lychee RGB-D image dataset (raw/augmented RGB images, depth maps, detection and maturity annotations) and the Python scripts for data augmentation, image similarity comparison, and annotation in the same GitHub repository.Dataset · publicchees, the non-augmented models
produced misclassifications with lower recognition and accuracy, whereas the augmented models
avoided these issues. Overall, the results demonstrate that the data augmentation method effectively
improves the comprehensive performance of the models.
5. Data Availability
The dataset is available at:https://github.com/SeiriosLab/Lychee. The Python scripts for data
augmentation, image similarity comparison, and annotation are available within the same
repository under the tree/main/script directory.Open asset ↗SeiriosLab/Lycheepdf-raw-page:13 lines:1-55Code / dataset availability confirmedarXiv · checked 13 Sept 2026
AI-driven crop health mapping systems offer substantial advantages over conventional monitoring approaches through accelerated data acquisition and cost reduction. However, widespread farmer adoption remains constrained by technical limitations in orthomosaic generation from sparse aerial imagery datasets. Traditional photogrammetric reconstruction requires 70-80\% inter-image overlap to establish sufficient feature correspondences for accurate geometric registration. AI-driven systems operating under resource-constrained conditions cannot consistently achieve these overlap thresholds, resulting in degraded reconstruction quality that undermines user confidence in autonomous monitoring technologies. In this paper, we present Ortho-Fuse, an optical flow-based framework that enables the generation of a reliable orthomosaic with reduced overlap requirements. Our approach employs intermediate flow estimation to synthesize transitional imagery between consecutive aerial frames, artificially augmenting feature correspondences for improved geometric reconstruction. Experimental validation demonstrates a 20\% reduction in minimum overlap requirements. We further analyze adoption barriers in precision agriculture to identify pathways for enhanced integration of AI-driven monitoring systems.
Why it matches plant phenotyping methods作物の健康状態マップ作成を目的とする航空画像のオルソモザイク生成法を開発し、重複率低減を実験検証しており、画像取得・再構成手法が中心である。
titleOrtho-Fuse: Orthomosaic Generation for Sparse High-Resolution Crop Health Maps Through Intermediate Optical Flow Estimation
Reproduction assets foundThe paper explicitly states that the authors' code and dataset (aerial imagery used for orthomosaic generation and crop health analysis) are publicly available at the project page https://rugvedkatole.github.io/OrthoFUSE/, which is an allowed URL. This qualifies as a paper-specific public asset containing the authors' Code · publicThe code and dataset are available at https://rugvedkatole.github.io/OrthoFUSE/Open asset ↗https://rugvedkatole.github.io/OrthoFUSE/lines:1-54Code / dataset availability confirmedarXiv · checked 15 Sept 2026
Street trees are vital to urban livability, providing ecological and social benefits. Establishing a detailed, accurate, and dynamically updated street tree inventory has become essential for optimizing these multifunctional assets within space-constrained urban environments. Given that traditional field surveys are time-consuming and labor-intensive, automated surveys utilizing Mobile Mapping Systems (MMS) offer a more efficient solution. However, existing MMS-acquired tree datasets are limited by small-scale scene, limited annotation, or single modality, restricting their utility for comprehensive analysis. To address these limitations, we introduce WHU-STree, a cross-city, richly annotated, and multi-modal urban street tree dataset. Collected across two distinct cities, WHU-STree integrates synchronized point clouds and high-resolution images, encompassing 21,007 annotated tree instances across 50 species and 2 morphological parameters. Leveraging the unique characteristics, WHU-STree concurrently supports over 10 tasks related to street tree inventory. We benchmark representative baselines for two key tasks--tree species classification and individual tree segmentation. Extensive experiments and in-depth analysis demonstrate the significant potential of multi-modal data fusion and underscore cross-domain applicability as a critical prerequisite for practical algorithm deployment. In particular, we identify key challenges and outline potential future works for fully exploiting WHU-STree, encompassing multi-modal fusion, multi-task collaboration, cross-domain generalization, spatial pattern learning, and Multi-modal Large Language Model for street tree asset management. The WHU-STree dataset is accessible at: https://github.com/WHU-USI3DV/WHU-STree.
Why it matches plant phenotyping methods樹木の個体セグメンテーションと形態パラメータを含むマルチモーダルデータセットを構築し、ベンチマークする研究であり、植物個体の状態・形態抽出手法が中心である。
abstractWHU-STree, a cross-city, richly annotated, and multi-modal urban street tree dataset.
Reproduction assets foundThe paper's core asset is the WHU-STree multi-modal street tree dataset (point clouds, panoramic images, 21,007 annotated tree instances, 50 species, height/DBH), which the authors state is publicly accessible via their GitHub organization WHU-USI3DV. The Zenodo DOIs in the reference list belong to cited prior datasetsDataset · publicticular, we
identify key challenges and outline potential future works for fully exploit-
ing WHU-STree, encompassing multi-modal fusion, multi-task collaboration,
cross-domain generalization, spatial pattern learning, and Multi-modal Large
Language Model for street tree asset management. The WHU-STree dataset
is accessible at: https://github.com/WHU-USI3DV /WHU-STree.
Keywords: Deep learning, Tree inventory, Individual tree segmentation,
Tree species classification, Multi-modal, Mobile mapping system
1. Introduction
Street trees, vital to urban ecosystems, provide ecological benefits (e.g.,
shade (Kumar et al., 2024), air purification (Grundstrém and Pleijel, 2014),
noise reductiOpen asset ↗WHU-STreepdf-raw-page:2 lines:1-35Code / dataset availability confirmedarXiv · checked 15 Sept 2026
Cotton is one of the most important natural fiber crops worldwide, yet harvesting remains limited by labor-intensive manual picking, low efficiency, and yield losses from missing the optimal harvest window. Accurate recognition of cotton bolls and their maturity is therefore essential for automation, yield estimation, and breeding research. We propose Cott-ADNet, a lightweight real-time detector tailored to cotton boll and flower recognition under complex field conditions. Building on YOLOv11n, Cott-ADNet enhances spatial representation and robustness through improved convolutional designs, while introducing two new modules: a NeLU-enhanced Global Attention Mechanism to better capture weak and low-contrast features, and a Dilated Receptive Field SPPF to expand receptive fields for more effective multi-scale context modeling at low computational cost. We curate a labeled dataset of 4,966 images, and release an external validation set of 1,216 field images to support future research. Experiments show that Cott-ADNet achieves 91.5% Precision, 89.8% Recall, 93.3% mAP50, 71.3% mAP, and 90.6% F1-Score with only 7.5 GFLOPs, maintaining stable performance under multi-scale and rotational variations. These results demonstrate Cott-ADNet as an accurate and efficient solution for in-field deployment, and thus provide a reliable basis for automated cotton harvesting and high-throughput phenotypic analysis. Code and dataset is available at https://github.com/SweefongWong/Cott-ADNet.
Why it matches plant phenotyping methods綿花の花・ボール認識を対象とする画像解析手法を開発し、データセット作成、外部検証、性能評価まで行っており、植物器官の表現型取得が中心である。
abstractWe propose Cott-ADNet, a lightweight real-time detector tailored to cotton boll and flower recognition under complex field conditions.
Reproduction assets foundThe paper explicitly states that its code and curated cotton boll/flower detection dataset (4,966 labeled images plus a 1,216-image external validation set) are publicly released at the authors' GitHub repository. The ultralytics repository is a generic third-party library, not a paper-specific asset.Code · publicy 7.5 GFLOPs, maintaining stable performance under multi-scale and rotational variations. These results demonstrate Cott-ADNet as an accurate and efficient solution for in-field deployment, and thus provide a reliable basis for automated cotton harvesting and high-throughput phenotypic analysis. Code and dataset is available at https://github.com/SweefongWong/Cott-ADNet .
† † footnotetext: ∗ * Corresponding author: cuij@wfu.edu
Index Terms :
cotton, cotton boll detection, lightweight object detection, rotational convolution
1 Introduction
Cotton is one of the most critical economic crops worldwide, accounting for nearly 35% of global natural fiber production. It underpins industries such as Open asset ↗SweefongWong/Cott-ADNetlines:1-57Code / dataset availability confirmedarXiv · checked 15 Sept 2026
StrawberryField / plotLiDAR / point cloudFlowerObject detectionPose / keypoint estimation
The small scale of urban farms and the commercial availability of low-cost robots (such as the FarmBot) that automate simple tending tasks enable an accessible platform for plant phenotyping. We have used a FarmBot with a custom camera end-effector to estimate strawberry plant flower pose (for robotic pollination) from acquired 3D point cloud models. We describe a novel algorithm that translates individual occupancy grids along orthogonal axes of a point cloud to obtain 2D images corresponding to the six viewpoints. For each image, 2D object detection models for flowers are used to identify 2D bounding boxes which can be converted into the 3D space to extract flower point clouds. Pose estimation is performed by fitting three shapes (superellipsoids, paraboloids and planes) to the flower point clouds and compared with manually labeled ground truth. Our method successfully finds approximately 80% of flowers scanned using our customized FarmBot platform and has a mean flower pose error of 7.7 degrees, which is sufficient for robotic pollination and rivals previous results. All code will be made available at https://github.com/harshmuriki/flowerPose.git.
Why it matches plant phenotyping methodsカスタムカメラ付きロボットによる3D花姿勢推定アルゴリズムとプラットフォームを開発・評価しており、花の姿勢という植物形質の取得が中心である。
abstractenable an accessible platform for plant phenotyping
Reproduction assets foundThe paper's flower pose estimation pipeline (translating occupancy grid, 2D/3D conversion, shape fitting) has an explicit authors' code deposit statement with a public GitHub URL, phrased as future availability ('will be made available'), so actionability is likely but not fully confirmed. No public dataset of the FarmCode · publiclower point clouds and compared with manually labeled ground truth. Our method successfully finds approximately 80% of flowers scanned using our customized FarmBot platform and has a mean flower pose error of 7.7 degrees, which is sufficient for robotic pollination and rivals previous results. All code will be made available at https://github.com/harshmuriki/flowerPose.git .
I Introduction
Urban farms [ 1 ] provide healthy food to local communities and can serve as platforms for education and sustainability. Unlike their rural counterparts, urban farms are usually small in scale and commercially available robotic systems such as the FarmBot [ 2 ] have been developed to help automate basic cuOpen asset ↗harshmuriki/flowerPoselines:1-53Code / dataset availability confirmedarXiv · checked 6 Sept 2026
Observer bias and inconsistencies in traditional plant phenotyping methods limit the accuracy and reproducibility of fine-grained plant analysis. To overcome these challenges, we developed TomatoMAP, a comprehensive dataset for Solanum lycopersicum using an Internet of Things (IoT) based imaging system with standardized data acquisition protocols. Our dataset contains 64,464 RGB images that capture 12 different plant poses from four camera elevation angles. Each image includes manually annotated bounding boxes for seven regions of interest (ROIs), including leaves, panicle, batch of flowers, batch of fruits, axillary shoot, shoot and whole plant area, along with 50 fine-grained growth stage classifications based on the BBCH scale. Additionally, we provide 3,616 high-resolution image subset with pixel-wise semantic and instance segmentation annotations for fine-grained phenotyping. We validated our dataset using a cascading model deep learning framework combining MobileNetv3 for classification, YOLOv11 for object detection, and MaskRCNN for segmentation. Through AI vs. Human analysis involving five domain experts, we demonstrate that the models trained on our dataset achieve accuracy and speed comparable to the experts. Cohen's Kappa and inter-rater agreement heatmap confirm the reliability of automated fine-grained phenotyping using our approach.
Why it matches plant phenotyping methods植物の多視点画像取得、アノテーション付きデータセット、深層学習による分類・検出・セグメンテーションを中心に開発・検証した、明確な植物フェノタイピング手法研究です。
abstractwe developed TomatoMAP, a comprehensive dataset for Solanum lycopersicum using an Internet of Things (IoT) based imaging system with standardized data acquisition protocols.
Reproduction assets foundThe paper's TomatoMAP dataset (images, annotations) is publicly deposited in e!DAL at IPK with an explicit DOI URL given in the Data Records section.Dataset · publicDataset is deposited in e!DAL (electronic data archive library) of IPK (Leibniz Institute of Plant Genetics
and Crop Plant Research): https://doi.ipk-gatersleben.de/DOI/10bb9f14-ce90-4747-836f-cf61dfb5eea1/Open asset ↗e!DAL · 10bb9f14-ce90-4747-836f-cf61dfb5eea1pdf-page:7 lines:1-73Code / dataset availability confirmedOpenAlex · arXiv · checked 15 Sept 2026
Quantitative descriptions of the complete canopy architecture are essential for accurately evaluating crop photosynthesis and yield performance to guide ideotype design. Although various sensing technologies have been developed for three-dimensional (3D) reconstruction of individual plants and canopies, they failed to obtain an accurate description of canopy architectures due to severe occlusion among complex canopy architectures. We proposed an effective method for 3D reconstruction of complex, dynamic population canopy architecture for rapeseed crops with a novel point cloud completion model. A complete point cloud generation framework was developed for automated annotation of the training dataset by distinguishing surface points from occluded points within canopies. The crop population point cloud completion network (CP-PCN) was then designed with a multi-resolution dynamic graph convolutional encoder (MRDG) and a point pyramid decoder (PPD) to predict occluded points. To further enhance feature extraction, a dynamic graph convolutional feature extractor (DGCFE) module was proposed to capture structural variations over the whole rapeseed growth period. The results demonstrated that CP-PCN achieved chamfer distance (CD) values of 3.35 cm -4.51 cm over four growth stages, outperforming the state-of-the-art transformer-based method (PoinTr). Ablation studies confirmed the effectiveness of the MRDG and DGCFE modules. Moreover, the validation experiment demonstrated that the silique efficiency index developed from CP-PCN improved the overall accuracy of rapeseed yield prediction by 11.2% compared to that of using incomplete point clouds. The CP-PCN pipeline has the potential to be extended to other crops, significantly advancing the quantitatively analysis of in-field population canopy architectures.
Why it matches plant phenotyping methods作物群落キャノピーの3D形態を復元する点群補完法を開発し、既存法との比較、アブレーション、収量予測への有効性検証まで行っており、植物フェノタイピング手法が研究の中心である。
abstractWe proposed an effective method for 3D reconstruction of complex, dynamic population canopy architecture for rapeseed crops with a novel point cloud completion model.
Reproduction assets foundThe paper's availability statement explicitly deposits all source code and test data (rapeseed canopy point cloud completion, CP-PCN) on GitHub at the allowed URL.Code · publicn Wang, Yi Feng,
Mengjie Gong and Guangyu Wu, for their participation in the experiments, and
to the Jiaxing Academy of Agricultural Sciences for their assistance with the
experimental data acquisition.
Availability of supporting data and source code
All source codes and test data involved in this study are available on
GitHub (https://github.com/Ziyue-Guo/RP-PCN.git).
Declaration of Competing Interest
The authors declare that they have no known competing financial interests
or personal relationships that could have appeared to influence the work
reported in this paper.
Contributions
Z. G. designed the study, conducted the experiments, and wrote the
manuscript. Y. S. contributed to the expeOpen asset ↗Ziyue-Guo/RP-PCNpdf-layout-page:42 lines:1-42Code / dataset availability confirmedarXiv · checked 15 Sept 2026
Accurate identification of individual plants from unmanned aerial vehicle (UAV) images is essential for advancing high-throughput phenotyping and supporting data-driven decision-making in plant breeding. This study presents MatchPlant, a modular, graphical user interface-supported, open-source Python pipeline for UAV-based single-plant detection and geospatial trait extraction. MatchPlant enables end-to-end workflows by integrating UAV image processing, user-guided annotation, Convolutional Neural Network model training for object detection, forward projection of bounding boxes onto an orthomosaic, and shapefile generation for spatial phenotypic analysis. In an early-season maize case study, MatchPlant achieved reliable detection performance (validation AP: 89.6%, test AP: 85.9%) and effectively projected bounding boxes, covering 89.8% of manually annotated boxes with 87.5% of projections achieving an Intersection over Union (IoU) greater than 0.5. Trait values extracted from predicted bounding instances showed high agreement with manual annotations (r = 0.87-0.97, IoU >= 0.4). Detection outputs were reused across time points to extract plant height and Normalized Difference Vegetation Index with minimal additional annotation, facilitating efficient temporal phenotyping. By combining modular design, reproducibility, and geospatial precision, MatchPlant offers a scalable framework for UAV-based plant-level analysis with broad applicability in agricultural and environmental monitoring.
Why it matches plant phenotyping methodsUAV画像から個体検出・地理空間的形質抽出を行うオープンソース基盤の開発と性能検証が中心であり、植物形質(草丈・NDVI)を抽出する再利用可能なワークフローを提供している。
abstractThis study presents MatchPlant, a modular, graphical user interface-supported, open-source Python pipeline for UAV-based single-plant detection and geospatial trait extraction.
Reproduction assets foundThe paper's MatchPlant pipeline code is publicly available on GitHub, and the maize case study training dataset and pre-trained model are publicly available on Zenodo.Dataset · publicinistration, Funding acquisition.
Declaration of Competing Interest
The authors declare that they have no known competing financial interests or personal relationships that could have appeared to influence the work reported in this paper.
Data availability
The public datasets supporting the case study are available on Zenodo at https://doi.org/10.5281/zenodo.14856123 (accessed on February 14, 2025). The source code and documentation for MatchPlant are available on GitHub at https://github.com/JacobWashburn-USDA/MatchPlant (accessed on February 14, 2025).
Acknowledgments
This research was supported in part by an appointment to the Agricultural Research Service (ARS) Research Participation PrOpen asset ↗Zenodo · 10.5281/zenodo.14856123lines:169-250Code / dataset availability confirmedarXiv · checked 6 Sept 2026
Plant phenotyping increasingly relies on (semi-)automated image-based analysis workflows to improve its accuracy and scalability. However, many existing solutions remain overly complex, difficult to reimplement and maintain, and pose high barriers for users without substantial computational expertise. To address these challenges, we introduce PhenoAssistant: a pioneering AI-driven system that streamlines plant phenotyping via intuitive natural language interaction. PhenoAssistant leverages a large language model to orchestrate a curated toolkit supporting tasks including automated phenotype extraction, data visualisation and automated model training. We validate PhenoAssistant through several representative case studies and a set of evaluation tasks. By significantly lowering technical hurdles, PhenoAssistant underscores the promise of AI-driven methodologies to democratising AI adoption in plant biology.
Why it matches plant phenotyping methods植物フェノタイピングの画像解析ワークフローを自然言語で自動化するシステムを開発し、ケーススタディと評価タスクで検証しているため、方法が中心である。
abstractwe introduce PhenoAssistant: a pioneering AI-driven system that streamlines plant phenotyping via intuitive natural language interaction.
Reproduction assets foundThe paper's authors release PhenoAssistant's code, chat logs, and evaluation results on GitHub, and the winter wheat nutrient-deficiency dataset used in Case Study 3 is publicly available on CodaLab. Case Study 1 demonstration data is request-only (Phenotiki), and Case Study 2 data is on Zenodo, which is not among the审Dataset · publicData for demonstrating Case Study 3 are publicly
available at https://codalab.lisn.upsaclay.fr/competitions/13833.Open asset ↗pdf-page:13 lines:1-47Code / dataset availability confirmedarXiv · checked 6 Sept 2026
ArabidopsisTomatoLeafRootSeed / grainSegmentationGrowth / time-series analysisTrackingGrowth / development / phenologyRoot system architecture
Plant developmental plasticity, particularly in root system architecture, is fundamental to understanding adaptability and agricultural sustainability. ChronoRoot 2.0 builds upon established low-cost hardware while significantly enhancing software capabilities and usability. The system employs nnUNet architecture for multi-class segmentation, demonstrating significant accuracy improvements while simultaneously tracking six distinct plant structures encompassing root, shoot, and seed components: main root, lateral roots, seed, hypocotyl, leaves, and petiole. This architecture enables easy retraining and incorporation of additional training data without requiring machine learning expertise. The platform introduces dual specialized graphical interfaces: a Standard Interface for detailed architectural analysis with novel gravitropic response parameters, and a Screening Interface enabling high-throughput analysis of multiple plants through automated tracking. Functional Principal Component Analysis integration enables discovery of novel phenotypic parameters through temporal pattern comparison. We demonstrate multi-species analysis, with Arabidopsis thaliana and Solanum lycopersicum, both morphologically distinct plant species. Three use cases in Arabidopsis thaliana and validation with tomato seedlings demonstrate enhanced capabilities: circadian growth pattern characterization, gravitropic response analysis in transgenic plants, and high-throughput etiolation screening across multiple genotypes.ChronoRoot 2.0 maintains the low-cost, modular hardware advantages of its predecessor while dramatically improving accessibility through intuitive graphical interfaces and expanded analytical capabilities. The open-source platform makes sophisticated temporal plant phenotyping more accessible to researchers without computational expertise.
Why it matches plant phenotyping methods植物の時系列画像から根・地上部・種子などの形態形質を抽出・追跡するオープンなAI基盤を開発し、精度向上、再学習、GUI、高スループット解析、検証まで扱っており、フェノタイピング手法が研究の中心です。
titleChronoRoot 2.0: An Open AI-Powered Platform for 2D Temporal Plant Phenotyping
Reproduction assets foundThe paper explicitly releases its full analysis source code (GitHub), the annotated infrared image dataset with multiclass segmentation masks (HuggingFace), demo phenotype video datasets, and a Docker image — all paper-specific, public, and actionable.Code · publicThe complete source code of ChronoRoot 2.0, including the implementation of all analysis methods described in this paper, is freely available under the GNU General Public License v3.0 at https://github.com/ChronoRoot/ChronoRoot2Open asset ↗ChronoRoot/ChronoRoot2lines:491-523Dataset · publicThe annotated image dataset used for training and validation contains 911 infrared images of Arabidopsis thaliana seedlings and 480 images of tomato with expert annotations for multiclass segmentation. This dataset is publicly available without restrictions at https://huggingface.co/datasets/ngaggion/ChronoRoot2Open asset ↗ngaggion/ChronoRoot2lines:491-523Code / dataset availability confirmedarXiv · checked 13 Sept 2026
Crop yield estimation is a relevant problem in agriculture, because an accurate yield estimate can support farmers' decisions on harvesting or precision intervention. Robots can help to automate this process. To do so, they need to be able to perceive the surrounding environment to identify target objects such as trees and plants. In this paper, we introduce a novel approach to address the problem of hierarchical panoptic segmentation of apple orchards on 3D data from different sensors. Our approach is able to simultaneously provide semantic segmentation, instance segmentation of trunks and fruits, and instance segmentation of trees (a trunk with its fruits). This allows us to identify relevant information such as individual plants, fruits, and trunks, and capture the relationship among them, such as precisely estimate the number of fruits associated to each tree in an orchard. To efficiently evaluate our approach for hierarchical panoptic segmentation, we provide a dataset designed specifically for this task. Our dataset is recorded in Bonn, Germany, in a real apple orchard with a variety of sensors, spanning from a terrestrial laser scanner to a RGB-D camera mounted on different robots platforms. The experiments show that our approach surpasses state-of-the-art approaches in 3D panoptic segmentation in the agricultural domain, while also providing full hierarchical panoptic segmentation. Our dataset is publicly available at https://www.ipb.uni-bonn.de/data/hops/. The open-source implementation of our approach is available at https://github.com/PRBonn/hapt3D.
Why it matches plant phenotyping methodsリンゴ樹・果実・幹を3Dセグメンテーションし、樹ごとの果実数を推定する手法と専用データセットを中心に開発・評価しており、植物の器官形態・収量関連形質の取得に該当する。
abstractwe introduce a novel approach to address the problem of hierarchical panoptic segmentation of apple orchards on 3D data from different sensors.
Reproduction assets foundThe paper introduces the HOPS dataset of annotated 3D apple orchard point clouds (TLS, UAV, UGV, SfM) for hierarchical panoptic segmentation, publicly available at the authors' IPB Bonn page, and releases the open-source implementation (hapt3D) on GitHub. Both are paper-specific, public, and actionable.Code · publicThe open-source implementation of our approach is available at https://github.com/PRBonn/hapt3D .Open asset ↗PRBonn/hapt3Dlines:1-59Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
Forest mapping provides critical observational data needed to understand the dynamics of forest environments. Notably, tree diameter at breast height (DBH) is a metric used to estimate forest biomass and carbon dioxide sequestration. Manual methods of forest mapping are labor intensive and time consuming, a bottleneck for large-scale mapping efforts. Automated mapping relies on acquiring dense forest reconstructions, typically in the form of point clouds. Terrestrial laser scanning (TLS) and mobile laser scanning (MLS) generate point clouds using expensive LiDAR sensing, and have been used successfully to estimate tree diameter. Neural radiance fields (NeRFs) are an emergent technology enabling photorealistic, vision-based reconstruction by training a neural network on a sparse set of input views. In this paper, we present a comparison of MLS and NeRF forest reconstructions for the purpose of trunk diameter estimation in a mixed-evergreen Redwood forest. In addition, we propose an improved DBH-estimation method using convex-hull modeling. Using this approach, we achieved 1.68 cm RMSE, which consistently outperformed standard cylinder modeling approaches. Our code contributions and forest datasets are freely available at https://github.com/harelab-ucsc/RedwoodNeRF.
Why it matches plant phenotyping methodsNeRFおよびMLSによる森林再構成から樹木DBHを推定し、凸包モデルによる推定法を提案・比較検証しているため、植物形質取得手法が中心です。
abstractIn this paper, we present a comparison of MLS and NeRF forest reconstructions for the purpose of trunk diameter estimation in a mixed-evergreen Redwood forest.
Reproduction assets foundThe authors explicitly state their code contributions and forest datasets (SLAM and NeRF reconstructions used for DBH estimation) are freely available in a public GitHub repository. The other URLs are a cited third-party tool (NeRFCapture) and a background reference (USDA aerial survey), neither of which is a paper-ownDataset · publicOur code contributions and forest datasets are freely available at https://github.com/harelab-ucsc/RedwoodNeRF .Open asset ↗harelab-ucsc/RedwoodNeRFlines:1-52Code / dataset availability confirmedarXiv · checked 13 Sept 2026
An early, non-invasive, and on-site detection of nutrient deficiencies is critical to enable timely actions to prevent major losses of crops caused by lack of nutrients. While acquiring labeled data is very expensive, collecting images from multiple views of a crop is straightforward. Despite its relevance for practical applications, unsupervised domain adaptation where multiple views are available for the labeled source domain as well as the unlabeled target domain is an unexplored research area. In this work, we thus propose an approach that leverages multiple camera views in the source and target domain for unsupervised domain adaptation. We evaluate the proposed approach on two nutrient deficiency datasets. The proposed method achieves state-of-the-art results on both datasets compared to other unsupervised domain adaptation methods. The dataset and source code are available at https://github.com/jh-yi/MV-Match.
Why it matches plant phenotyping methods植物の栄養欠乏状態を画像から推定するマルチビュー・ドメイン適応手法を提案し、2つのデータセットで評価しており、表現型取得・推定手法が中心です。
abstractwe thus propose an approach that leverages multiple camera views in the source and target domain for unsupervised domain adaptation.
Reproduction assets foundThe paper's authors explicitly state that the MiPlo nutrient-deficiency image datasets and the MV-Match source code are publicly available at the authors' GitHub repository, which matches an allowed URL.Dataset · publicThe dataset and source code are available at https://github.com/jh-yi/MV-Match .Open asset ↗jh-yi/MV-Matchlines:1-71Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
AppleCitrusMangoPeachPearPlumField / plotLiDAR / point cloudRGB / grayscaleFruit
We introduce FruitNeRF, a unified novel fruit counting framework that leverages state-of-the-art view synthesis methods to count any fruit type directly in 3D. Our framework takes an unordered set of posed images captured by a monocular camera and segments fruit in each image. To make our system independent of the fruit type, we employ a foundation model that generates binary segmentation masks for any fruit. Utilizing both modalities, RGB and semantic, we train a semantic neural radiance field. Through uniform volume sampling of the implicit Fruit Field, we obtain fruit-only point clouds. By applying cascaded clustering on the extracted point cloud, our approach achieves precise fruit count.The use of neural radiance fields provides significant advantages over conventional methods such as object tracking or optical flow, as the counting itself is lifted into 3D. Our method prevents double counting fruit and avoids counting irrelevant fruit.We evaluate our methodology using both real-world and synthetic datasets. The real-world dataset consists of three apple trees with manually counted ground truths, a benchmark apple dataset with one row and ground truth fruit location, while the synthetic dataset comprises various fruit types including apple, plum, lemon, pear, peach, and mango.Additionally, we assess the performance of fruit counting using the foundation model compared to a U-Net.
Why it matches plant phenotyping methods果実を対象に、画像・NeRF・点群クラスタリングを組み合わせて3D果実数を推定する手法を開発し、実データおよび合成データで評価しているため、植物表現型取得法が中心である。
abstractWe introduce FruitNeRF, a unified novel fruit counting framework that leverages state-of-the-art view synthesis methods to count any fruit type directly in 3D.
Reproduction assets foundThe paper's real-world apple tree image dataset with manual ground-truth counts and synthetic Blender fruit tree data are publicly released via the project website, and the FruitNeRF analysis code is open-source on GitHub. The Zenodo DOI refers to the third-party BlenderNeRF plugin (cited tool), not a paper-specific.Dataset · publicThe data has been made publicly available, and visualizations can be accessed on the project website.Open asset ↗lines:183-221Code · publicFruitNeRF code: https://github.com/meyerls/FruitNeRF has been made open-source.Open asset ↗meyerls/FruitNeRFlines:74-108Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
Potato yield is an important metric for farmers to further optimize their cultivation practices. Potato yield can be estimated on a harvester using an RGB-D camera that can estimate the three-dimensional (3D) volume of individual potato tubers. A challenge, however, is that the 3D shape derived from RGB-D images is only partially completed, underestimating the actual volume. To address this issue, we developed a 3D shape completion network, called CoRe++, which can complete the 3D shape from RGB-D images. CoRe++ is a deep learning network that consists of a convolutional encoder and a decoder. The encoder compresses RGB-D images into latent vectors that are used by the decoder to complete the 3D shape using the deep signed distance field network (DeepSDF). To evaluate our CoRe++ network, we collected partial and complete 3D point clouds of 339 potato tubers on an operational harvester in Japan. On the 1425 RGB-D images in the test set (representing 51 unique potato tubers), our network achieved a completion accuracy of 2.8 mm on average. For volumetric estimation, the root mean squared error (RMSE) was 22.6 ml, and this was better than the RMSE of the linear regression (31.1 ml) and the base model (36.9 ml). We found that the RMSE can be further reduced to 18.2 ml when performing the 3D shape completion in the center of the RGB-D image. With an average 3D shape completion time of 10 milliseconds per tuber, we can conclude that CoRe++ is both fast and accurate enough to be implemented on an operational harvester for high-throughput potato yield estimation. CoRe++'s high-throughput and accurate processing allows it to be applied to other tuber, fruit and vegetable crops, thereby enabling versatile, accurate and real-time yield monitoring in precision agriculture. Our code, network weights and dataset are publicly available at https://github.com/UTokyo-FieldPhenomics-Lab/corepp.git.
Why it matches plant phenotyping methodsRGB-D画像からジャガイモ塊茎の3D形状を補完し、体積・収量を推定する手法を開発・検証しており、植物フェノタイピング手法が研究の中心である。
abstractwe developed a 3D shape completion network, called CoRe++, which can complete the 3D shape from RGB-D images.
Reproduction assets foundThe paper's abstract explicitly states that the authors' code, network weights, and the potato tuber RGB-D/3D point cloud dataset are publicly available at the authors' GitHub repository (UTokyo-FieldPhenomics-Lab/corepp), which is a paper-specific, public, actionable asset for the CoRe++ phenotyping analysis.Code · publicOur code, network weights and dataset are publicly available at https://github.com/UTokyo-FieldPhenomics-Lab/corepp.git .Open asset ↗UTokyo-FieldPhenomics-Lab/corepplines:1-93Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
Creation of new annotated public datasets is crucial in helping advances in 3D computer vision and machine learning meet their full potential for automatic interpretation of 3D plant models. Despite the proliferation of deep neural network architectures for segmentation and phenotyping of 3D plant models in the last decade, the amount of data, and diversity in terms of species and data acquisition modalities are far from sufficient for evaluation of such tools for their generalization ability. To contribute to closing this gap, we introduce PLANesT-3D; a new annotated dataset of 3D color point clouds of plants. PLANesT-3D is composed of 34 point cloud models representing 34 real plants from three different plant species: \textit{Capsicum annuum}, \textit{Rosa kordana}, and \textit{Ribes rubrum}. Both semantic labels in terms of "leaf" and "stem", and organ instance labels were manually annotated for the full point clouds. PLANesT-3D introduces diversity to existing datasets by adding point clouds of two new species and providing 3D data acquired with the low-cost SfM/MVS technique as opposed to laser scanning or expensive setups. Point clouds reconstructed with SfM/MVS modality exhibit challenges such as missing data, variable density, and illumination variations. As an additional contribution, SP-LSCnet, a novel semantic segmentation method that is a combination of unsupervised superpoint extraction and a 3D point-based deep learning approach is introduced and evaluated on the new dataset. The advantages of SP-LSCnet over other deep learning methods are its modular structure and increased interpretability. Two existing deep neural network architectures, PointNet++ and RoseSegNet, were also tested on the point clouds of PLANesT-3D for semantic segmentation.
Why it matches plant phenotyping methods3D植物点群の注釈付きデータセットを構築し、植物器官のセマンティック・インスタンス分割手法を開発・評価しており、植物フェノタイピング手法が中心である。
abstractwe introduce PLANesT-3D; a new annotated dataset of 3D color point clouds of plants.
Reproduction assets foundThe paper introduces PLANesT-3D, an annotated 3D plant point cloud dataset, and SP-LSCnet segmentation code, both explicitly stated as publicly available at the authors' Aperta record and GitHub repository.Dataset · publicThe PLANesT-3D dataset is publicly available at https://aperta.ulakbim.gov.tr/record/286354 and https://github.com/visionlab-ogu/PLANesT-3D/tree/main/dataOpen asset ↗aperta.ulakbim.gov.tr · 286354lines:83-145Dataset · publicThe 2D color images for all the 34 plants together with their estimated camera poses and parameters are also open to the public to provide input data for recent 3D reconstruction techniques 3 3
3
The data is available at https://github.com/visionlab-ogu/PLANesT-3D/tree/main/data .Open asset ↗github.com/visionlab-ogu/PLANesT-3Dlines:494-505Code · publicThe code for SP-LSCnet is available at https://github.com/visionlab-ogu/PLANesT-3DOpen asset ↗github.com/visionlab-ogu/PLANesT-3Dlines:146-154Code / dataset availability confirmedarXiv · checked 14 Sept 2026
As the world population is expected to reach 10 billion by 2050, our agricultural production system needs to double its productivity despite a decline of human workforce in the agricultural sector. Autonomous robotic systems are one promising pathway to increase productivity by taking over labor-intensive manual tasks like fruit picking. To be effective, such systems need to monitor and interact with plants and fruits precisely, which is challenging due to the cluttered nature of agricultural environments causing, for example, strong occlusions. Thus, being able to estimate the complete 3D shapes of objects in presence of occlusions is crucial for automating operations such as fruit harvesting. In this paper, we propose the first publicly available 3D shape completion dataset for agricultural vision systems. We provide an RGB-D dataset for estimating the 3D shape of fruits. Specifically, our dataset contains RGB-D frames of single sweet peppers in lab conditions but also in a commercial greenhouse. For each fruit, we additionally collected high-precision point clouds that we use as ground truth. For acquiring the ground truth shape, we developed a measuring process that allows us to record data of real sweet pepper plants, both in the lab and in the greenhouse with high precision, and determine the shape of the sensed fruits. We release our dataset, consisting of almost 7,000 RGB-D frames belonging to more than 100 different fruits. We provide segmented RGB-D frames, with camera intrinsics to easily obtain colored point clouds, together with the corresponding high-precision, occlusion-free point clouds obtained with a high-precision laser scanner. We additionally enable evaluation of shape completion approaches on a hidden test set through a public challenge on a benchmark server.
Why it matches plant phenotyping methods果実の3D形状という植物器官の形態形質を対象に、RGB-D画像・高精度点群・評価用ベンチマークを構築しており、形状取得と推定の方法論が中心である。
abstractWe provide an RGB-D dataset for estimating the 3D shape of fruits.
Reproduction assets found保存済みの本文根拠を更新済みルールで再検証し、公開資産1件を確認しました。Code · publicOur development toolkit including a data loader is available at:
https://github.com/PRBonn/shape_completion_toolkit for handling the dataset and computing metrics.Open asset ↗PRBonn/shape_completion_toolkitlines:55-81Code / dataset availability confirmedarXiv · checked 15 Sept 2026
High-throughput phenotyping refers to the non-destructive and efficient evaluation of plant phenotypes. In recent years, it has been coupled with machine learning in order to improve the process of phenotyping plants by increasing efficiency in handling large datasets and developing methods for the extraction of specific traits. Previous studies have developed methods to advance these challenges through the application of deep neural networks in tandem with automated cameras; however, the datasets being studied often excluded physical labels. In this study, we used a dataset provided by Oak Ridge National Laboratory with 1,672 images of Populus Trichocarpa with white labels displaying treatment (control or drought), block, row, position, and genotype. Optical character recognition (OCR) was used to read these labels on the plants, image segmentation techniques in conjunction with machine learning algorithms were used for morphological classifications, machine learning models were used to predict treatment based on those classifications, and analyzed encoded EXIF tags were used for the purpose of finding leaf size and correlations between phenotypes. We found that our OCR model had an accuracy of 94.31% for non-null text extractions, allowing for the information to be accurately placed in a spreadsheet. Our classification models identified leaf shape, color, and level of brown splotches with an average accuracy of 62.82%, and plant treatment with an accuracy of 60.08%. Finally, we identified a few crucial pieces of information absent from the EXIF tags that prevented the assessment of the leaf size. There was also missing information that prevented the assessment of correlations between phenotypes and conditions. However, future studies could improve upon this to allow for the assessment of these features.
Why it matches plant phenotyping methods植物画像からラベル情報を読み取り、画像分割・機械学習で葉形、色、斑点などの形態形質を抽出・分類する手法が研究の中心であり、植物フェノタイピング手法の開発・適用に該当する。
abstractimage segmentation techniques in conjunction with machine learning algorithms were used for morphological classifications
Reproduction assets foundThe paper's authors explicitly state that all analysis code (OCR label reading, leaf segmentation, morphology classification, treatment prediction) is publicly available under the MIT License on their GitHub repository. The underlying ORNL image dataset is not stated to be publicly available, so only the code asset is.Code · publicSince a pre-trained segmentation model (the SAM) was used in this study, researchers could attempt to build segmentation models fine-tuned to only recognize leaves, which could increase model efficiency and provide more consistent results.
6 Code Availability
All code is publicly available under the MIT License on GitHub here: https://github.com/vivaansinghvi07/smoky-mountain-data-comp .
Acknowledgements
We thank Dr. Ty Frazier at Oak Ridge National Laboratory for his helpful suggestions and mentoring throughout this project.
References
Arya et al. (2022)
Arya, S., Sandhu, K.S.,
Singh, J., Kumar, S.,
2022.
Deep learning: As the new frontier in high-throughput
plant phenotyping.
Euphytica 218Open asset ↗vivaansinghvi07/smoky-mountain-data-complines:272-401Code / dataset availability confirmedarXiv · checked 13 Sept 2026
The process of estimating and counting tree density using only a single aerial or satellite image is a difficult task in the fields of photogrammetry and remote sensing. However, it plays a crucial role in the management of forests. The huge variety of trees in varied topography severely hinders tree counting models to perform well. The purpose of this paper is to propose a framework that is learnt from the source domain with sufficient labeled trees and is adapted to the target domain with only a limited number of labeled trees. Our method, termed as AdaTreeFormer, contains one shared encoder with a hierarchical feature extraction scheme to extract robust features from the source and target domains. It also consists of three subnets: two for extracting self-domain attention maps from source and target domains respectively and one for extracting cross-domain attention maps. For the latter, an attention-to-adapt mechanism is introduced to distill relevant information from different domains while generating tree density maps; a hierarchical cross-domain feature alignment scheme is proposed that progressively aligns the features from the source and target domains. We also adopt adversarial learning into the framework to further reduce the gap between source and target domains. Our AdaTreeFormer is evaluated on six designed domain adaptation tasks using three tree counting datasets, \ie Jiangsu, Yosemite, and London. Experimental results show that AdaTreeFormer significantly surpasses the state of the art, \eg in the cross domain from the Yosemite to Jiangsu dataset, it achieves a reduction of 15.9 points in terms of the absolute counting errors and an increase of 10.8\% in the accuracy of the detected trees' locations. The codes and datasets are available at https://github.com/HAAClassic/AdaTreeFormer.
Why it matches plant phenotyping methods単一の航空・衛星画像から樹木数・密度を推定する画像解析手法を開発し、複数データセットとドメイン適応タスクで評価しており、植物形質の取得が中心である。
abstractThe purpose of this paper is to propose a framework that is learnt from the source domain with sufficient labeled trees and is adapted to the target domain with only a limited number of labeled trees.
Reproduction assets foundThe paper publicly releases its AdaTreeFormer code and datasets, and evaluates on three publicly available tree-counting image/annotation datasets (Jiangsu, London, Yosemite) with explicit GitHub availability statements.Code · publicThe codes and datasets are available at https://github.com/HAAClassic/AdaTreeFormer .Open asset ↗HAAClassic/AdaTreeFormerlines:1-70Dataset · publicThis dataset encompasses 24 satellite images taken by the GaofenII satellite with a ground sample distance (GSD) of 0.8m (available at https://github.com/sddpltwanqiu/TreeCountNet/tree/main).Open asset ↗sddpltwanqiu/TreeCountNetlines:201-252Dataset · publicThis dataset consists of high-resolution images captured at 0.2m GSD from London, United Kingdom for training and testing (available at https://github.com/HAAClassic/TreeFormer/tree/main).Open asset ↗HAAClassic/TreeFormerlines:201-252Dataset · publicThe study area for this dataset revolves around Yosemite National Park, located in California, United States of America (available at https://github.com/nightonion/yosemite-tree-dataset ).Open asset ↗nightonion/yosemite-tree-datasetlines:201-252Code / dataset availability confirmedarXiv · OpenAlex · checked 15 Sept 2026
Automatic tree density estimation and counting using single aerial and satellite images is a challenging task in photogrammetry and remote sensing, yet has an important role in forest management. In this paper, we propose the first semisupervised transformer-based framework for tree counting which reduces the expensive tree annotations for remote sensing images. Our method, termed as TreeFormer, first develops a pyramid tree representation module based on transformer blocks to extract multi-scale features during the encoding stage. Contextual attention-based feature fusion and tree density regressor modules are further designed to utilize the robust features from the encoder to estimate tree density maps in the decoder. Moreover, we propose a pyramid learning strategy that includes local tree density consistency and local tree count ranking losses to utilize unlabeled images into the training process. Finally, the tree counter token is introduced to regulate the network by computing the global tree counts for both labeled and unlabeled images. Our model was evaluated on two benchmark tree counting datasets, Jiangsu, and Yosemite, as well as a new dataset, KCL-London, created by ourselves. Our TreeFormer outperforms the state of the art semi-supervised methods under the same setting and exceeds the fully-supervised methods using the same number of labeled images. The codes and datasets are available at https://github.com/HAAClassic/TreeFormer.
Why it matches plant phenotyping methods樹木の個体数・密度という植物状態を航空・衛星画像から推定する画像解析手法を開発し、複数データセットで評価しているため、植物フェノタイピング手法が中心である。
abstractAutomatic tree density estimation and counting using single aerial and satellite images is a challenging task
Reproduction assets foundThe paper's authors publicly release their analysis code and the KCL-London tree counting dataset via GitHub, and the paper's annotation workflow directly uses the public London Datastore local-authority-maintained trees dataset for tree locations.Code · publichmark tree counting datasets, Jiangsu, and Yosemite, as well as a new dataset, KCL-London, created by ourselves. Our TreeFormer outperforms the state of the art semi-supervised methods under the same setting and exceeds the fully-supervised methods using the same number of labeled images. The codes and datasets are available at https://github.com/HAAClassic/TreeFormer .
Index Terms:
Tree counting, semi-supervised model, transformer, pyramid learning strategy, remote sensing.
I Introduction
Trees are the pulse of the earth and are vital organisms in maintaining the ecological functioning and health of the planet [ 1 ] . Tree counting using high-resolution images is useful in various fields suOpen asset ↗HAAClassic/TreeFormerlines:1-71Dataset · publicmages are gathered and stitched together from Google Maps at 0.2 m ground sampling distance (GSD). The gathered images are divided into images with 1024 × 1024 pixels.
To aid the identification of tree locations and numbers of selected images, we employed the accessible tree locations of London in London Datastore website 1 1
1
https://data.london.gov.uk/dataset/local-authority-maintained-trees .
Although these data show the locations and species information for over 880,000 of London’s trees, the data mainly contains information on trees in the main streets and does not cover trees that are dense between houses or parks. We manually annotated the latter.
To this end, Global Mapper as geograOpen asset ↗lines:124-145Code / dataset availability confirmedarXiv · OpenAlex · checked 13 Sept 2026
Artificial intelligence applications enable farmers to optimize crop growth and production while reducing costs and environmental impact. Computer vision-based algorithms in particular, are commonly used for fruit segmentation, enabling in-depth analysis of the harvest quality and accurate yield estimation. In this paper, we propose TomatoDIFF, a novel diffusion-based model for semantic segmentation of on-plant tomatoes. When evaluated against other competitive methods, our model demonstrates state-of-the-art (SOTA) performance, even in challenging environments with highly occluded fruits. Additionally, we introduce Tomatopia, a new, large and challenging dataset of greenhouse tomatoes. The dataset comprises high-resolution RGB-D images and pixel-level annotations of the fruits.
Why it matches plant phenotyping methods植物上のトマト果実を画像からセグメンテーションする手法を開発・比較し、RGB-D画像と画素アノテーションのデータセットも提供しており、植物器官の状態・位置推定に関わる方法が中心である。
abstractwe propose TomatoDIFF, a novel diffusion-based model for semantic segmentation of on-plant tomatoes
Reproduction assets foundThe paper introduces TomatoDIFF and the Tomatopia dataset, with explicit public availability of source code and dataset at the authors' GitHub repository. It also trains/evaluates on the public Kaggle 'Tomato dataset' (andrewmvd/tomato-detection), which is a paper-specific public image dataset used directly in the phenCode · publicThe source code of TomatoDIFF and Tomatopia are available at https://github.com/MIvanovska/TomatoDIFF .Open asset ↗MIvanovska/TomatoDIFFlines:1-44Code / dataset availability confirmedarXiv · checked 14 Sept 2026
Multispectral / hyperspectralWhole plant / canopy / plot / field
The diversity of terrestrial vascular plants plays a key role in maintaining the stability and productivity of ecosystems. Airborne hyperspectral imaging has shown promise for measuring plant diversity remotely, but to operationalise these efforts over large regions we need to advance satellite-based alternatives. The advanced spectral and spatial specification of the recently launched DESIS (the DLR Earth Sensing Imaging Spectrometer) instrument provides a unique opportunity to test the potential for monitoring plant species diversity with spaceborne hyperspectral data. This study provides a quantitative assessment on the ability of DESIS hyperspectral data for predicting plant species richness in two different habitat types in southeast Australia. Spectral features were first extracted from the DESIS spectra, then regressed against on-ground estimates of plant species richness, with a two-fold cross validation scheme to assess the predictive performance. We tested and compared the effectiveness of Principal Component Analysis (PCA), Canonical Correlation Analysis (CCA), and Partial Least Squares analysis (PLS) for feature extraction, and Kernel Ridge Regression (KRR), Gaussian Process Regression (GPR), and Random Forest Regression (RFR) for species richness prediction. The best prediction results were $r=0.76$ and $\text{RMSE}=5.89$ for the Southern Tablelands region, and $r=0.68$ and $\text{RMSE}=5.95$ for the Snowy Mountains region. Relative importance analysis for the DESIS spectral bands showed that the red-edge, red, and blue spectral regions were more important for predicting plant species richness than the green bands and the near-infrared bands beyond red-edge. We also found that the DESIS hyperspectral data performed better than Sentinel-2 multispectral data in the prediction of plant species richness.
Why it matches plant phenotyping methodsDESISハイパースペクトルデータから植物種数を推定する特徴抽出・回帰手法を比較し、交差検証で性能評価しており、植物フェノタイピング手法が中心である。
abstractThis study provides a quantitative assessment on the ability of DESIS hyperspectral data for predicting plant species richness in two different habitat types in southeast Australia.
Reproduction assets found保存済みの本文根拠を更新済みルールで再検証し、公開資産1件を確認しました。Dataset · publicFor on-ground measures of vascular plant species richness, we obtained plant community survey data from the NSW BioNet Vegetation Information System database [ Government, 2019 ] .Open asset ↗NSW BioNet Vegetation Information Systemlines:75-98Code / dataset availability confirmedarXiv · OpenAlex · checked 15 Sept 2026
In intensively managed forests in Europe, where forests are divided into stands of small size and may show heterogeneity within stands, a high spatial resolution (10 - 20 meters) is arguably needed to capture the differences in canopy height. In this work, we developed a deep learning model based on multi-stream remote sensing measurements to create a high-resolution canopy height map over the "Landes de Gascogne" forest in France, a large maritime pine plantation of 13,000 km$^2$ with flat terrain and intensive management. This area is characterized by even-aged and mono-specific stands, of a typical length of a few hundred meters, harvested every 35 to 50 years. Our deep learning U-Net model uses multi-band images from Sentinel-1 and Sentinel-2 with composite time averages as input to predict tree height derived from GEDI waveforms. The evaluation is performed with external validation data from forest inventory plots and a stereo 3D reconstruction model based on Skysat imagery available at specific locations. We trained seven different U-net models based on a combination of Sentinel-1 and Sentinel-2 bands to evaluate the importance of each instrument in the dominant height retrieval. The model outputs allow us to generate a 10 m resolution canopy height map of the whole "Landes de Gascogne" forest area for 2020 with a mean absolute error of 2.02 m on the Test dataset. The best predictions were obtained using all available satellite layers from Sentinel-1 and Sentinel-2 but using only one satellite source also provided good predictions. For all validation datasets in coniferous forests, our model showed better metrics than previous canopy height models available in the same region.
Why it matches plant phenotyping methodsSentinel/GEDI等のリモートセンシング画像から樹冠高を推定する深層学習手法を開発し、外部データで検証しているため、植物形質取得法が中心である。
abstractwe developed a deep learning model based on multi-stream remote sensing measurements to create a high-resolution canopy height map
Reproduction assets foundThe paper's primary phenotyping-relevant input is the GEDI L2A canopy height dataset (526,449 footprints over the Landes forest, 2020), explicitly downloaded from NASA's EarthDataSearch. This is a public, paper-specific sensor dataset directly used for the study's canopy height measurements and model training. No code,Dataset · publicwater bodies (Beck et al., 2020). Indeed, these
surfaces mirror the transmitted waveforms that have a pulse width of ~ 15 ns which
corresponds to a ~ 2.25 m wide waveform (Dubayah et al., 2020).
In total, 526,449 footprints from the GEDIv002 L2A product (Dubayah et al., 2021) were
downloaded from NASA’s EarthDataSearch website
(https://search.earthdata.nasa.gov/search) for this study, covering the entire area of interest
for 2020. Due to atmospheric perturbations, some waveforms could not be used to give
information on the vertical forest structure. Therefore, several filtering criteria were applied to
remove unusable waveforms: (1) When the quality_flag provided in the GEDI data was set toOpen asset ↗GEDIv002 L2Apdf-raw-page:6 lines:1-45Code / dataset availability confirmedarXiv · OpenAlex · checked 15 Sept 2026
In this paper, we present a method for creating high-quality 3D models of sorghum panicles for phenotyping in breeding experiments. This is achieved with a novel reconstruction approach that uses seeds as semantic landmarks in both 2D and 3D. To evaluate the performance, we develop a new metric for assessing the quality of reconstructed point clouds without having a ground-truth point cloud. Finally, a counting method is presented where the density of seed centers in the 3D model allows 2D counts from multiple views to be effectively combined into a whole-panicle count. We demonstrate that using this method to estimate seed count and weight for sorghum outperforms count extrapolation from 2D images, an approach used in most state of the art methods for seeds and grains of comparable size.
Why it matches plant phenotyping methodsソルガム穂の3D再構成と種子計数という植物形質取得手法を開発し、再構成品質評価指標と種子数・重量推定を検証しており、フェノタイピング手法が中心である。
abstractwe present a method for creating high-quality 3D models of sorghum panicles for phenotyping in breeding experiments
Reproduction assets foundThe paper's authors publicly release their sorghum panicle stereo-image dataset (camera poses, human-labeled seed segmentations, panicle weights, seed counts) via the CMU AIIRA resources page, which is an allowed URL.Dataset · publicection, some unremoved husks were counted as seeds by the counting machine despite manual efforts to separate seeds from husks. We expect the effect on the ground truth to be small. The stereo images, camera poses, human-labeled seed segmentations, panicle weights, and human-counted seed counts can be found in our dataset 3 3
3
https://labs.ri.cmu.edu/aiira/resources/ .
Figure 8: (a) 100 sorghum panicles from 10 different sorghum species. (b) Our data collection system, a stereo camera attached to the UR5 robot arm. (c) Seeds were manually stripped and (d) counted using a seed counting machine.
IV-B 3D Reconstruction Quality
We assess the effectiveness of our approach with ablation tests usiOpen asset ↗lines:141-165Code / dataset availability confirmedarXiv · OpenAlex · checked 14 Sept 2026
We propose a novel hybrid cable-based robot with manipulator and camera for high-accuracy, medium-throughput plant monitoring in a vertical hydroponic farm and, as an example application, demonstrate non-destructive plant mass estimation. Plant monitoring with high temporal and spatial resolution is important to both farmers and researchers to detect anomalies and develop predictive models for plant growth. The availability of high-quality, off-the-shelf structure-from-motion (SfM) and photogrammetry packages has enabled a vibrant community of roboticists to apply computer vision for non-destructive plant monitoring. While existing approaches tend to focus on either high-throughput (e.g. satellite, unmanned aerial vehicle (UAV), vehicle-mounted, conveyor-belt imagery) or high-accuracy/robustness to occlusions (e.g. turn-table scanner or robot arm), we propose a middle-ground that achieves high accuracy with a medium-throughput, highly automated robot. Our design pairs the workspace scalability of a cable-driven parallel robot (CDPR) with the dexterity of a 4 degree-of-freedom (DoF) robot arm to autonomously image many plants from a variety of viewpoints. We describe our robot design and demonstrate it experimentally by collecting daily photographs of 54 plants from 64 viewpoints each. We show that our approach can produce scientifically useful measurements, operate fully autonomously after initial calibration, and produce better reconstructions and plant property estimates than those of over-canopy methods (e.g. UAV). As example applications, we show that our system can successfully estimate plant mass with a Mean Absolute Error (MAE) of 0.586g and, when used to perform hypothesis testing on the relationship between mass and age, produces p-values comparable to ground-truth data (p=0.0020 and p=0.0016, respectively).
Why it matches plant phenotyping methods植物の多視点画像取得、SfM再構成、質量推定を中核とするロボット型フェノタイピング手法の開発・実証であり、単なる生物学的測定ではない。
abstractWe describe our robot design and demonstrate it experimentally by collecting daily photographs of 54 plants from 64 viewpoints each.
Obtaining 3D sensor data of complete plants or plant parts (e.g., the crop or fruit) is difficult due to their complex structure and a high degree of occlusion. However, especially for the estimation of the position and size of fruits, it is necessary to avoid occlusions as much as possible and acquire sensor information of the relevant parts. Global viewpoint planners exist that suggest a series of viewpoints to cover the regions of interest up to a certain degree, but they usually prioritize global coverage and do not emphasize the avoidance of local occlusions. On the other hand, there are approaches that aim at avoiding local occlusions, but they cannot be used in larger environments since they only reach a local maximum of coverage. In this paper, we therefore propose to combine a local, gradient-based method with global viewpoint planning to enable local occlusion avoidance while still being able to cover large areas. Our simulated experiments with a robotic arm equipped with a camera array as well as an RGB-D camera show that this combination leads to a significantly increased coverage of the regions of interest compared to just applying global coverage planning.
Why it matches plant phenotyping methods果実の位置・サイズ推定に必要な3Dセンサデータ取得を対象に、局所遮蔽回避と大域的視点計画を組み合わせる視点計画法を開発・評価しており、植物表現型取得が中心である。
abstractespecially for the estimation of the position and size of fruits, it is necessary to avoid occlusions as much as possible and acquire sensor information of the relevant parts
Reproduction assets foundThe paper's authors explicitly state that the source code of their combined local/global viewpoint planning system (used for fruit ROI coverage experiments) is publicly available on GitHub. OctoMap is a generic third-party library and is excluded.Code · publicThe source code of our system is available on GitHub 1 1
1
https://github.com/Eruvae/roi_viewpoint_planner .Open asset ↗Eruvae/roi_viewpoint_plannerlines:1-105Code / dataset availability confirmedarXiv · checked 15 Sept 2026
Field / plotRGB / grayscaleMultispectral / hyperspectralThermalWhole plant / canopy / plot / field
A core objective of the TERRA-REF project was to generate an open-access reference dataset for the evaluation of sensing technologies to study plants under field conditions. The TERRA-REF program deployed a suite of high-resolution, cutting edge technology sensors on a gantry system with the aim of scanning 1 hectare (10$^4$) at around 1 mm$^2$ spatial resolution multiple times per week. The system contains co-located sensors including a stereo-pair RGB camera, a thermal imager, a laser scanner to capture 3D structure, and two hyperspectral cameras covering wavelengths of 300-2500nm. This sensor data is provided alongside over sixty types of traditional plant phenotype measurements that can be used to train new machine learning models. Associated weather and environmental measurements, information about agronomic management and experimental design, and the genomic sequences of hundreds of plant varieties have been collected and are available alongside the sensor and plant phenotype data. Over the course of four years and ten growing seasons, the TERRA-REF system generated over 1 PB of sensor data and almost 45 million files. The subset that has been released to the public domain accounts for two seasons and about half of the total data volume. This provides an unprecedented opportunity for investigations far beyond the core biological scope of the project. The focus of this paper is to provide the Computer Vision and Machine Learning communities an overview of the available data and some potential applications of this one of a kind data.
Why it matches plant phenotyping methods植物の高解像度マルチセンサーデータと植物表現型データを含む公開ベンチマーク/データセットを紹介し、コンピュータビジョンでの利用を主目的とするため、フェノタイピング手法・基盤として中心的です。
abstractgenerate an open-access reference dataset for the evaluation of sensing technologies to study plants under field conditions
Reproduction assets foundThe paper describes the TERRA-REF public domain release of plant phenotyping sensor data (RGB, thermal, laser scanner, hyperspectral, PSII) plus derived phenotypes, and explicitly points to public code repositories for the processing pipeline (terraref GitHub, PhytoOracle, AgPipeline) and a data access portal. All are,Dataset · publicprocessing, reviewing, curating, describing, and hosting the data.
Instead, we focused on an initial public release and plan to make new datasets available based on need.
Access to unpublished data can be requested from the authors, and as data are curated they will be added to subsequent versions of the public domain release ( https://terraref.org/data/access-data ).
In addition to hosting an archival copy of data on Dryad [ 16 ] , the
documentation includes instructions for browsing and accessing these
data through a variety of online portals. These portals provide access
to web user interfaces as well as databases, APIs, and R and Python
clients. In some cases it will be easier to acceOpen asset ↗lines:234-317Code · publicapproach described by Li et al . [ 18 ] .
Herritt et al . [ 14 , 13 ] demonstrate and provide software used in analysis of a sequence of images that capture plant fluorescence response to a pulse of light.
Most of the algorithms used to generate data products have not been published as papers but are made available on GitHub ( https://github.com/terraref ); code
used to release the data publication in 2020 is available on Zenodo [ 25 , 15 , 10 , 6 , 4 , 19 , 8 , 7 , 5 , 9 , 17 ] .
Pipeline development continues to support ongoing use of the field scanner as well as more general applications in plant sensing pipelines.
Recent advances have improved pipeline scalability and modulOpen asset ↗terrareflines:193-233Code · publiclant sensing pipelines.
Recent advances have improved pipeline scalability and modularity by adopting workflow tools and making use of heterogeneous computing environments.
The TERRA-REF computing pipeline has been adapted and extended for continuing use with the Field Scanner with the new name ”PhytoOracle” and is available at https://github.com/LyonsLab/PhytoOracle . Related work generalizing the pipeline for other phenomics applications has been released under the name ”AgPipeline” https://github.com/agpipeline with applications to aerial imaging described by Schnaufer et al . [ 22 ] .
All of these software are made available with permissive open source licenses on GitHub to enable accesOpen asset ↗PhytoOraclelines:193-233Code · publicnvironments.
The TERRA-REF computing pipeline has been adapted and extended for continuing use with the Field Scanner with the new name ”PhytoOracle” and is available at https://github.com/LyonsLab/PhytoOracle . Related work generalizing the pipeline for other phenomics applications has been released under the name ”AgPipeline” https://github.com/agpipeline with applications to aerial imaging described by Schnaufer et al . [ 22 ] .
All of these software are made available with permissive open source licenses on GitHub to enable access and community development.
Figure 4: Summary of public sensor datasets from Seasons 4 and 6. Each dot represents the dates for which a particular daOpen asset ↗agpipelinelines:193-233Code / dataset availability confirmedarXiv · checked 13 Sept 2026
Retrieval of vegetation properties from satellite and airborne optical data usually takes place after atmospheric correction, yet it is also possible to develop retrieval algorithms directly from top-of-atmosphere (TOA) radiance data. One of the key vegetation variables that can be retrieved from at-sensor TOA radiance data is the leaf area index (LAI) if algorithms account for variability in the atmosphere. We demonstrate the feasibility of LAI retrieval from Sentinel-2 (S2) TOA radiance data (L1C product) in a hybrid machine learning framework. To achieve this, the coupled leaf-canopy-atmosphere radiative transfer models PROSAIL-6S were used to simulate a look-up table (LUT) of TOA radiance data and associated input variables. This LUT was then used to train the Bayesian machine learning algorithms Gaussian processes regression (GPR) and variational heteroscedastic GPR (VHGPR). PROSAIL simulations were also used to train GPR and VHGPR models for LAI retrieval from S2 images at bottom-of-atmosphere (BOA) level (L2A product) for comparison purposes. The VHGPR models led to consistent LAI maps at BOA and TOA scale. We demonstrated that hybrid LAI retrieval algorithms can be developed from TOA radiance data given a cloud-free sky, thus without the need for atmospheric correction.
Why it matches plant phenotyping methodsSentinel-2のTOA放射輝度からLAIを推定する機械学習アルゴリズムを開発・比較しており、植物形質の取得手法が研究の中心である。
abstractWe demonstrate the feasibility of LAI retrieval from Sentinel-2 (S2) TOA radiance data (L1C product) in a hybrid machine learning framework.
Reproduction assets foundThe paper's hybrid LAI retrieval (GPR/VHGPR) was developed within the authors' ALG-ARTMO software framework, and code snippets/demos for GPR and VHGPR are publicly available from the authors' UV-ES soft regression page. Both are explicitly stated as freely downloadable in the supplied text. No paper-specific phenotype/Code · publicCode snippets and demos for both GPR, VHGPR and other machine learning regression algorithms is available from https://isp.uv.es/soft_regression.html .Open asset ↗isp.uv.es/soft_regression.htmllines:485-521Code / dataset availability confirmedarXiv · checked 13 Sept 2026
Modern agricultural applications require knowledge about the position and size of fruits on plants. However, occlusions from leaves typically make obtaining this information difficult. We present a novel viewpoint planning approach that builds up an octree of plants with labeled regions of interest (ROIs), i.e., fruits. Our method uses this octree to sample viewpoint candidates that increase the information around the fruit regions and evaluates them using a heuristic utility function that takes into account the expected information gain. Our system automatically switches between ROI targeted sampling and exploration sampling, which considers general frontier voxels, depending on the estimated utility. When the plants have been sufficiently covered with the RGB-D sensor, our system clusters the ROI voxels and estimates the position and size of the detected fruits. We evaluated our approach in simulated scenarios and compared the resulting fruit estimations with the ground truth. The results demonstrate that our combined approach outperforms a sampling method that does not explicitly consider the ROIs to generate viewpoints in terms of the number of discovered ROI cells. Furthermore, we show the real-world applicability by testing our framework on a robotic arm equipped with an RGB-D camera installed on an automated pipe-rail trolley in a capsicum glasshouse.
Why it matches plant phenotyping methods果実の位置・サイズという植物器官形質をRGB-Dセンサで取得・推定する視点計画法を開発し、シミュレーションと実環境で検証しており、フェノタイピング手法が中心である。
abstractWe present a novel viewpoint planning approach that builds up an octree of plants with labeled regions of interest (ROIs), i.e., fruits.
Reproduction assets foundThe paper's viewpoint-planning system source code and the simulated capsicum plant environments used in the experiments are publicly available on GitHub. OctoMap is a generic third-party library, not a paper-specific asset.Code · publicThe source code of our system is available on GitHub 1 1
1
https://github.com/Eruvae/roi_viewpoint_planner .Open asset ↗Eruvae/roi_viewpoint_plannerlines:73-113Code / dataset availability confirmedarXiv · checked 9 Sept 2026
Supervised learning is often used to count objects in images, but for counting small, densely located objects, the required image annotations are burdensome to collect. Counting plant organs for image-based plant phenotyping falls within this category. Object counting in plant images is further challenged by having plant image datasets with significant domain shift due to different experimental conditions, e.g. applying an annotated dataset of indoor plant images for use on outdoor images, or on a different plant species. In this paper, we propose a domain-adversarial learning approach for domain adaptation of density map estimation for the purposes of object counting. The approach does not assume perfectly aligned distributions between the source and target datasets, which makes it more broadly applicable within general object counting and plant organ counting tasks. Evaluation on two diverse object counting tasks (wheat spikelets, leaves) demonstrates consistent performance on the target datasets across different classes of domain shift: from indoor-to-outdoor images and from species-to-species adaptation.
Why it matches plant phenotyping methods植物器官数を画像から推定するドメイン適応・密度マップ推定法を開発し、コムギ小穂と葉の計数で評価しており、表現型取得手法が中心である。
abstractCounting plant organs for image-based plant phenotyping falls within this category.
Reproduction assets foundThe paper provides two paper-specific public assets: the authors' implementation code for the domain-adversarial counting model on GitHub, and the authors' newly created GWHD dot annotations deposited on figshare. Other URLs are cited prior-work datasets, not paper-specific assets.Code · publicAll experiments were performed on a GeForce RTX 2070 GPU with 8GB memory using the Pytorch framework. The implementation is available at: https://github.com/p2irc/UDA4POCOpen asset ↗p2irc/UDA4POClines:81-104Dataset · publicTo evaluate our method, we created dot annotations for 67 images from the GWHD which are used as ground truth. These annotations are made publicly available at https://doi.org/10.6084/m9.figshare.12652973.v2 .Open asset ↗10.6084/m9.figshare.12652973.v2lines:105-155Code / dataset availability confirmedarXiv · checked 15 Sept 2026
Providing an accurate evaluation of palm tree plantation in a large region can bring meaningful impacts in both economic and ecological aspects. However, the enormous spatial scale and the variety of geological features across regions has made it a grand challenge with limited solutions based on manual human monitoring efforts. Although deep learning based algorithms have demonstrated potential in forming an automated approach in recent years, the labelling efforts needed for covering different features in different regions largely constrain its effectiveness in large-scale problems. In this paper, we propose a novel domain adaptive oil palm tree detection method, i.e., a Multi-level Attention Domain Adaptation Network (MADAN) to reap cross-regional oil palm tree counting and detection. MADAN consists of 4 procedures: First, we adopted a batch-instance normalization network (BIN) based feature extractor for improving the generalization ability of the model, integrating batch normalization and instance normalization. Second, we embedded a multi-level attention mechanism (MLA) into our architecture for enhancing the transferability, including a feature level attention and an entropy level attention. Then we designed a minimum entropy regularization (MER) to increase the confidence of the classifier predictions through assigning the entropy level attention value to the entropy penalty. Finally, we employed a sliding window-based prediction and an IOU based post-processing approach to attain the final detection results. We conducted comprehensive ablation experiments using three different satellite images of large-scale oil palm plantation area with six transfer tasks. MADAN improves the detection accuracy by 14.98% in terms of average F1-score compared with the Baseline method (without DA), and performs 3.55%-14.49% better than existing domain adaptation methods.
Why it matches plant phenotyping methods油ヤシ個体の計数・検出という植物形態/個体数形質を衛星画像から推定する手法を開発し、アブレーション実験と既存手法比較で検証しており、フェノタイピング手法が中心である。
abstractwe propose a novel domain adaptive oil palm tree detection method, i.e., a Multi-level Attention Domain Adaptation Network (MADAN) to reap cross-regional oil palm tree counting and detection.
Reproduction assets foundThe authors explicitly state that their code and datasets (satellite images and annotations used for oil palm tree detection) are publicly available on GitHub.Code · publiche oil palm tree detection performance across different
remotely sensed images acquired from different sensors, regions and dates, without using labeled samples in the
target region. Our MADAN is proposed for enhancing both the generalization capacity and the transferability of
our model. Our codes and datasets are available on https://github.com/rs-dl/MADAN. The major contributions of
our work are as follows:
(1) We propose an adaptive object detector named MADAN for oil palm tree counting and detection across different
satellite images, which is the first work for large-scale domain adaptive tree crown detection using multi-source and
multi-temporal remote sensing images.
(2) WeOpen asset ↗rs-dl/MADANpdf-raw-page:6 lines:1-22Code / dataset availability confirmedarXiv · checked 14 Sept 2026
Deep learning models have been successfully deployed for a diverse array of image-based plant phenotyping applications including disease detection and classification. However, successful deployment of supervised deep learning models requires large amount of labeled data, which is a significant challenge in plant science (and most biological) domains due to the inherent complexity. Specifically, data annotation is costly, laborious, time consuming and needs domain expertise for phenotyping tasks, especially for diseases. To overcome this challenge, active learning algorithms have been proposed that reduce the amount of labeling needed by deep learning models to achieve good predictive performance. Active learning methods adaptively select samples to annotate using an acquisition function to achieve maximum (classification) performance under a fixed labeling budget. We report the performance of four different active learning methods, (1) Deep Bayesian Active Learning (DBAL), (2) Entropy, (3) Least Confidence, and (4) Coreset, with conventional random sampling-based annotation for two different image-based classification datasets. The first image dataset consists of soybean [Glycine max L. (Merr.)] leaves belonging to eight different soybean stresses and a healthy class, and the second consists of nine different weed species from the field. For a fixed labeling budget, we observed that the classification performance of deep learning models with active learning-based acquisition strategies is better than random sampling-based acquisition for both datasets. The integration of active learning strategies for data annotation can help mitigate labelling challenges in the plant sciences applications particularly where deep domain knowledge is required.
Why it matches plant phenotyping methods植物画像フェノタイピングにおける能動学習手法を比較評価しており、ラベル付け削減と分類性能が中心的な方法論的貢献である。
titleHow useful is Active Learning for Image-based Plant Phenotyping?
Reproduction assets found保存済みの本文根拠を更新済みルールで再検証し、公開資産1件を確認しました。Code · publicAll the codes for the active learning approaches described in this work are available for the community at https://github.com/koushik-n/Active-Learning-Plant-Phenotyping.Open asset ↗koushik-n/Active-Learning-Plant-Phenotypinglines:108-129Code / dataset availability confirmedarXiv · checked 15 Sept 2026
Deep neural networks have shown excellent performances in many real-world applications. Unfortunately, they may show "Clever Hans"-like behavior -- making use of confounding factors within datasets -- to achieve high performance. In this work, we introduce the novel learning setting of "explanatory interactive learning" (XIL) and illustrate its benefits on a plant phenotyping research task. XIL adds the scientist into the training loop such that she interactively revises the original model via providing feedback on its explanations. Our experimental results demonstrate that XIL can help avoiding Clever Hans moments in machine learning and encourages (or discourages, if appropriate) trust into the underlying model.
Why it matches plant phenotyping methods植物フェノタイピング課題を対象に、説明への研究者フィードバックを学習ループへ組み込む新しい機械学習手法を提案・実証しており、フェノタイピング解析手法が中心である。
abstractIn this work, we introduce the novel learning setting of "explanatory interactive learning" (XIL) and illustrate its benefits on a plant phenotyping research task.
Reproduction assets foundThe paper's plant phenotyping RGB/hyperspectral dataset is publicly deposited on TU Datalib, and the authors' analysis code (runnable Code Ocean capsule with pre-trained models reproducing figures/results) plus the user study materials are publicly available on GitHub. Generic benchmarks (Fashion-MNIST, PASCAL VOC) areCode · publiclable at
https://github.com/zalandoresearch/fashion-mnist . The PASCAL VOC2007 dataset is available at http://host.robots.ox.ac.uk/pascal/VOC/voc2007/ .
The RGB and hyperspectral data that support the findings of this study are available at https://tudatalib.ulb.tu-darmstadt.de/handle/tudatalib/2278.4 and in the code repository https://codeocean.com/capsule/4559958/tree .
The user study is available at https://github.com/ml-research/xil/tree/master/Trust_Study .
Code availability
The code and a fully runnable capsule to reproduce the figures and results of this article, including pre-trained models, can be found at https://codeocean.com/capsule/4559958/tree .
Statement of ethical complianceOpen asset ↗codeocean · capsule/4559958lines:369-382Code / dataset availability confirmedarXiv · OpenAlex · checked 14 Sept 2026
Images are used frequently in plant phenotyping to capture measurements. This chapter offers a repeatable method for capturing two-dimensional measurements of plant parts in field or laboratory settings using a variety of camera styles (cellular phone, DSLR), with the addition of a printed calibration pattern. The method is based on calibrating the camera using information available from the EXIF tags from the image, as well as visual information from the pattern. Code is provided to implement the method, as well as a dataset for testing. We include steps to verify protocol correctness by imaging an artifact. The use of this protocol for two-dimensional plant phenotyping will allow data capture from different cameras and environments, with comparison on the same physical scale. We abbreviate this method as CASS, for CAmera aS Scanner. Code and data is available at http://doi.org/10.5281/zenodo.3677473.
Why it matches plant phenotyping methods植物部位の2次元形質をカメラで測定する手法を開発し、校正・検証手順、コード、テストデータを提供しており、植物フェノタイピング手法が中心である。
abstractThis chapter offers a repeatable method for capturing two-dimensional measurements of plant parts in field or laboratory settings using a variety of camera styles (cellular phone, DSLR), with the addition of a printed calibration pattern.
Reproduction assets found保存済みの本文根拠を更新済みルールで再検証し、公開資産1件を確認しました。Dataset · publicThe code and test datasets are provided in [ 16 ] . Within [ 16 ] , are some example data and two programs:
aruco-pattern-write and camera-as-scanner . To prepare for the experiments, download the example data and install the code (C++ code as well as a Docker image are provided).Open asset ↗lines:54-78Code / dataset availability confirmedOpenAlex · arXiv · checked 10 Sept 2026
A looming question that must be solved before robotic plant phenotyping capabilities can have significant impact to crop improvement programs is scalability. High Throughput Phenotyping (HTP) uses robotic technologies to analyze crops in order to determine species with favorable traits, however, the current practices rely on exhaustive coverage and data collection from the entire crop field being monitored under the breeding experiment. This works well in relatively small agricultural fields but can not be scaled to the larger ones, thus limiting the progress of genetics research. In this work, we propose an active learning algorithm to enable an autonomous system to collect the most informative samples in order to accurately learn the distribution of phenotypes in the field with the help of a Gaussian Process model. We demonstrate the superior performance of our proposed algorithm compared to the current practices on sorghum phenotype data collection.
Why it matches plant phenotyping methods植物表現型データ収集を効率化する能動学習・ガウス過程アルゴリズムを開発し、ソルガムの表現型データで既存手法と比較評価しており、表現型取得ワークフローが中心である。
abstractIn this work, we propose an active learning algorithm to enable an autonomous system to collect the most informative samples in order to accurately learn the distribution of phenotypes in the field with the help of a Gaussian Process model.
Reproduction assets foundThe authors explicitly open-sourced their code repository, simulation environment, and the sorghum phenotype dataset used in this paper, with a public GitHub URL.Code · publicWe have open-sourced our code repository, simulation environment and the sorghum dataset 1 1
1
Our github repository can be found at https://github.com/sumitsk/algp.git for the research community to carry out further work in this direction.Open asset ↗sumitsk/algp · sumitsk/algplines:55-75Dataset · publicOur simulation environment, code repository and the sorghum dataset are open-sourced and can be found at https://github.com/sumitsk/algp.git .Open asset ↗sumitsk/algp · sumitsk/algplines:180-200Code / dataset availability confirmedOpenAlex · arXiv · checked 10 Sept 2026
Automated segmentation of individual leaves of a plant in an image is a prerequisite to measure more complex phenotypic traits in high-throughput phenotyping. Applying state-of-the-art machine learning approaches to tackle leaf instance segmentation requires a large amount of manually annotated training data. Currently, the benchmark datasets for leaf segmentation contain only a few hundred labeled training images. In this paper, we propose a framework for leaf instance segmentation by augmenting real plant datasets with generated synthetic images of plants inspired by domain randomisation. We train a state-of-the-art deep learning segmentation architecture (Mask-RCNN) with a combination of real and synthetic images of Arabidopsis plants. Our proposed approach achieves 90% leaf segmentation score on the A1 test set outperforming the-state-of-the-art approaches for the CVPPP Leaf Segmentation Challenge (LSC). Our approach also achieves 81% mean performance over all five test datasets.
Why it matches plant phenotyping methods植物画像から個葉を自動セグメンテーションする手法を開発・評価しており、植物表現型抽出のための方法が中心的です。
abstractAutomated segmentation of individual leaves of a plant in an image is a prerequisite to measure more complex phenotypic traits in high-throughput phenotyping.
Reproduction assets foundThe paper's authors publicly released their generated synthetic Arabidopsis dataset (10,000 top-down images with 2D segmentation labels) used for training their leaf segmentation models, hosted on the CSIRO robotics databases page. The Matterport Mask_RCNN repository is a generic third-party library, and the CodaLab L5Dataset · publicOur generated synthetic dataset is publicly available at 3 3
3
https://research.csiro.au/robotics/databases . The synthetic dataset contains 10,000 top down images of synthetic Arabidopsis plants and their corresponding 2D segmentation labels.Open asset ↗lines:242-290Code / dataset availability confirmedOpenAlex · arXiv · checked 10 Sept 2026
We propose a new and, arguably, a very simple reduction of instance segmentation to semantic segmentation. This reduction allows to train feed-forward non-recurrent deep instance segmentation systems in an end-to-end fashion using architectures that have been proposed for semantic segmentation. Our approach proceeds by introducing a fixed number of labels (colors) and then dynamically assigning object instances to those labels during training (coloring). A standard semantic segmentation objective is then used to train a network that can color previously unseen images. At test time, individual object instances can be recovered from the output of the trained convolutional network using simple connected component analysis. In the experimental validation, the coloring approach is shown to be capable of solving diverse instance segmentation tasks arising in autonomous driving (the Cityscapes benchmark), plant phenotyping (the CVPPP leaf segmentation challenge), and high-throughput microscopy image analysis. The source code is publicly available: https://github.com/kulikovv/DeepColoring.
Why it matches plant phenotyping methodsインスタンスセグメンテーション手法の開発と実験検証が中心で、植物フェノタイピングの葉セグメンテーション課題に明示的に適用されている。
abstractWe propose a new and, arguably, a very simple reduction of instance segmentation to semantic segmentation.
Reproduction assets foundThe paper applies its Deep Coloring instance segmentation method to plant phenotyping (CVPPP leaf segmentation) and states its PyTorch implementation is publicly available on GitHub, enabling reproduction of the phenotyping analysis.Code · publicThe source code is publicly available: https://github.com/kulikovv/DeepColoring.Open asset ↗kulikovv/DeepColoringpdf-page:1 lines:1-64Code / dataset availability confirmedarXiv · OpenAlex · checked 15 Sept 2026
In recent years, there has been an increasing interest in image-based plant phenotyping, applying state-of-the-art machine learning approaches to tackle challenging problems, such as leaf segmentation (a multi-instance problem) and counting. Most of these algorithms need labelled data to learn a model for the task at hand. Despite the recent release of a few plant phenotyping datasets, large annotated plant image datasets for the purpose of training deep learning algorithms are lacking. One common approach to alleviate the lack of training data is dataset augmentation. Herein, we propose an alternative solution to dataset augmentation for plant phenotyping, creating artificial images of plants using generative neural networks. We propose the Arabidopsis Rosette Image Generator (through) Adversarial Network: a deep convolutional network that is able to generate synthetic rosette-shaped plants, inspired by DCGAN (a recent adversarial network model using convolutional layers). Specifically, we trained the network using A1, A2, and A4 of the CVPPP 2017 LCC dataset, containing Arabidopsis Thaliana plants. We show that our model is able to generate realistic 128x128 colour images of plants. We train our network conditioning on leaf count, such that it is possible to generate plants with a given number of leaves suitable, among others, for training regression based models. We propose a new Ax dataset of artificial plants images, obtained by our ARIGAN. We evaluate this new dataset using a state-of-the-art leaf counting algorithm, showing that the testing error is reduced when Ax is used as part of the training data.
Why it matches plant phenotyping methods植物表現型解析用の合成画像生成ネットワークを開発し、葉数を条件付けたデータセットを作成・評価しており、表現型取得・解析ワークフローの技術的貢献が中心である。
abstractWe propose a new Ax dataset of artificial plants images, obtained by our ARIGAN.
Reproduction assets foundThe paper's authors publicly released the Ax dataset of 57 synthetic Arabidopsis plant images generated by ARIGAN (with leaf-count annotations in a CSV), which directly reproduces the paper's phenotyping data contribution. The CVPPP 2017 LCC dataset is the training input but is cited prior work, not a paper-specific.Dataset · publicr quantitative experiments show that the extension of the training dataset with the images in Ax improved the testing error and reduced overfitting. We run a 4-fold cross validation experiment on A4 dataset. Evaluation metrics of our experiments are reported in Table 1 . Our synthetic dataset Ax is available to download at \url http://www.valeriogiuffrida.academy/ax.
Acknowledgements
This work was supported by The Alan Turing Institute under the EPSRC grant EP/N510129/1, and also by the BBSRC grant BB/P023487/1.
References
[1]
F. Bastien, P. Lamblin, R. Pascanu, J. Bergstra, I. J. Goodfellow, A. Bergeron,
N. Bouchard, and Y. Bengio.
Theano: new features and speed improvements.
Deep LearninOpen asset ↗Axlines:101-153Code / dataset availability confirmedarXiv · checked 10 Sept 2026
In this paper, we investigate the problem of counting rosette leaves from an RGB image, an important task in plant phenotyping. We propose a data-driven approach for this task generalized over different plant species and imaging setups. To accomplish this task, we use state-of-the-art deep learning architectures: a deconvolutional network for initial segmentation and a convolutional network for leaf counting. Evaluation is performed on the leaf counting challenge dataset at CVPPP-2017. Despite the small number of training samples in this dataset, as compared to typical deep learning image sets, we obtain satisfactory performance on segmenting leaves from the background as a whole and counting the number of leaves using simple data augmentation strategies. Comparative analysis is provided against methods evaluated on the previous competition datasets. Our framework achieves mean and standard deviation of absolute count difference of 1.62 and 2.30 averaged over all five test datasets.
Why it matches plant phenotyping methodsロゼット葉の画像から葉数を推定する深層学習手法を提案し、セグメンテーションと葉数カウントをデータセットで評価しており、植物表現型取得手法が中心である。
abstractcounting rosette leaves from an RGB image, an important task in plant phenotyping
Reproduction assets foundThe paper's authors explicitly state their leaf counting/segmentation code is publicly available on GitHub, and the CVPPP2017 Leaf Counting Challenge dataset used for all experiments is publicly hosted at plant-phenotyping.org.Code · publicCode is publicly available here. 1 1Open asset ↗lines:215-305Code / dataset availability confirmedarXiv · checked 14 Sept 2026
Phenomics is an emerging branch of modern biology that uses high throughput phenotyping tools to capture multiple environmental and phenotypic traits, often at massive spatial and temporal scales. The resulting high dimensional data represent a treasure trove of information for providing an in-depth understanding of how multiple factors interact and contribute to the overall growth and behavior of different genotypes. However, computational tools that can parse through such complex data and aid in extracting plausible hypotheses are currently lacking. In this paper, we present Hyppo-X, a new algorithmic approach to visually explore complex phenomics data and in the process characterize the role of environment on phenotypic traits. We model the problem as one of unsupervised structure discovery, and use emerging principles from algebraic topology and graph theory for discovering higher-order structures of complex phenomics data. We present an open source software which has interactive visualization capabilities to facilitate data navigation and hypothesis formulation. We test and evaluate Hyppo-X on two real-world plant (maize) data sets. Our results demonstrate the ability of our approach to delineate divergent subpopulation-level behavior. Notably, our approach shows how environmental factors could influence phenotypic behavior, and how that effect varies across different genotypes and different time scales. To the best of our knowledge, this effort provides one of the first approaches to systematically formalize the problem of hypothesis extraction for phenomics data. Considering the infancy of the phenomics field, tools that help users explore complex data and extract plausible hypotheses in a data-guided manner will be critical to future advancements in the use of such data.
Why it matches plant phenotyping methods植物フェノミクスデータを探索・解析するアルゴリズムとオープンソースソフトウェアを開発し、植物データセットで評価しているため、フェノタイピング解析手法が中心である。
abstractIn this paper, we present Hyppo-X, a new algorithmic approach to visually explore complex phenomics data
Reproduction assets found保存済みの本文根拠を更新済みルールで再検証し、公開資産1件を確認しました。Code · publicThe tool is available as open source in the GitHub repository [ 18 ] .Open asset ↗lines:185-261