← Papers

Unverified paper record

On the Minimum Dataset Requirements for Fine-Tuning an Object Detector for Arable Crop Plant Counting: A Case Study on Maize Seedlings

Remote Sensing · 25 Jun 2025 · 10.3390/rs17132190

Abstract

Object detection is essential for precision agriculture applications like automated plant counting, but the minimum dataset requirements for effective model deployment remain poorly understood for arable crop seedling detection on orthomosaics. This study investigated how much annotated data is required to achieve standard counting accuracy (R2 = 0.85) for maize seedlings across different object detection approaches. We systematically evaluated traditional deep learning models requiring many training examples (YOLOv5, YOLOv8, YOLO11, RT-DETR), newer approaches requiring few examples (CD-ViTO), and methods requiring zero labeled examples (OWLv2) using drone-captured orthomosaic RGB imagery. We also implemented a handcrafted computer graphics algorithm as baseline. Models were tested with varying training sources (in-domain vs. out-of-distribution data), training dataset sizes (10–150 images), and annotation quality levels (10–100%). Our results demonstrate that no model trained on out-of-distribution data achieved acceptable performance, regardless of dataset size. In contrast, models trained on in-domain data reached the benchmark with as few as 60–130 annotated images, depending on architecture. Transformer-based models (RT-DETR) required significantly fewer samples (60) than CNN-based models (110–130), though they showed different tolerances to annotation quality reduction. Models maintained acceptable performance with only 65–90% of original annotation quality. Despite recent advances, neither few-shot nor zero-shot approaches met minimum performance requirements for precision agriculture deployment. These findings provide practical guidance for developing maize seedling detection systems, demonstrating that successful deployment requires in-domain training data, with minimum dataset requirements varying by model architecture.

Plant phenotyping relevance

トウモロコシ幼苗の個体数という植物形質を画像から推定する物体検出手法について、複数モデル、データ量、アノテーション品質を系統的に比較・評価しており、手法の性能検証が中心である。

abstractThis study investigated how much annotated data is required to achieve standard counting accuracy (R2 = 0.85) for maize seedlings across different object detection approaches.
abstractWe systematically evaluated traditional deep learning models requiring many training examples (YOLOv5, YOLOv8, YOLO11, RT-DETR), newer approaches requiring few examples (CD-ViTO), and methods requiring zero labeled examples (OWLv2) using drone-captured orthomosaic RGB imagery.
abstractThese findings provide practical guidance for developing maize seedling detection systems, demonstrating that successful deployment requires in-domain training data, with minimum dataset requirements varying by model architecture.

Code and data availability

The paper's Data Availability Statement provides two paper-specific public assets: the authors' handcrafted-method analysis code on a GitHub gist and the ID (in-distribution) annotation datasets created for this study on Zenodo. Both have explicit availability language and public URLs.

Codepublic

The code for the handcrafted methods used in this study is available at https://gist.github.com/SamueleBumbaca/4a227bbe7b78d6be3424899c16c60bb4 (accessed on 20 June 2025).

Open resource ↗gist.github.com/SamueleBumbaca · pdf-page:23 lines:1-52
Datasetpublic

The datasets created during this study (ID datasets) are available at the Zenodo repository https://doi.org/10.5281/zenodo.15235602 (accessed on 20 June 2025)

Open resource ↗Zenodo · 10.5281/zenodo.15235602 · pdf-page:23 lines:1-52

This is an automatically classified, unverified record. Curator approval is required before any resource enters the Catalog.