← Papers

Unverified paper record

Citizen crowds and experts: observer variability in image-based plant phenotyping

Plant Methods · 9 Feb 2018 · 10.1186/s13007-018-0278-7

Abstract

BACKGROUND: Image-based plant phenotyping has become a powerful tool in unravelling genotype-environment interactions. The utilization of image analysis and machine learning have become paramount in extracting data stemming from phenotyping experiments. Yet we rely on observer (a human expert) input to perform the phenotyping process. We assume such input to be a 'gold-standard' and use it to evaluate software and algorithms and to train learning-based algorithms. However, we should consider whether any variability among experienced and non-experienced (including plain citizens) observers exists. Here we design a study that measures such variability in an annotation task of an integer-quantifiable phenotype: the leaf count. RESULTS: to measure intra- and inter-observer variability in a controlled study using specially designed annotation tools but also citizens using a distributed citizen-powered web-based platform. In the controlled study observers counted leaves by looking at top-view images, which were taken with low and high resolution optics. We assessed whether the utilization of tools specifically designed for this task can help to reduce such variability. We found that the presence of tools helps to reduce intra-observer variability, and that although intra- and inter-observer variability is present it does not have any effect on longitudinal leaf count trend statistical assessments. We compared the variability of citizen provided annotations (from the web-based platform) and found that plain citizens can provide statistically accurate leaf counts. We also compared a recent machine-learning based leaf counting algorithm and found that while close in performance it is still not within inter-observer variability. CONCLUSIONS: While expertise of the observer plays a role, if sufficient statistical power is present, a collection of non-experienced users and even citizens can be included in image-based phenotyping annotation tasks as long they are suitably designed. We hope with these findings that we can re-evaluate the expectations that we have from automated algorithms: as long as they perform within observer variability they can be considered a suitable alternative. In addition, we hope to invigorate an interest in introducing suitably designed tasks on citizen powered platforms not only to obtain useful information (for research) but to help engage the public in this societal important problem.

Plant phenotyping relevance

画像ベース植物フェノタイピングにおける葉数アノテーションの観察者間・内変動を、専用ツール、市民参加型基盤、機械学習アルゴリズムと比較検証しており、測定手法の技術的妥当性評価が中心である。

abstractHere we design a study that measures such variability in an annotation task of an integer-quantifiable phenotype: the leaf count.
abstractto measure intra- and inter-observer variability in a controlled study using specially designed annotation tools but also citizens using a distributed citizen-powered web-based platform.
abstractWe also compared a recent machine-learning based leaf counting algorithm and found that while close in performance it is still not within inter-observer variability.

Code and data availability

The article states that the Arabidopsis image dataset used for the leaf-counting observer-variability study is publicly available at the plant-phenotyping.org datasets page, and the citizen-science annotations were collected via the authors' public Zooniverse 'Leaf Targeting' project. No author analysis code or trained

Datasetpublic

Availability of data and materials The image dataset used in this article is available at http://www.plant-phenotyping.org/datasets .

Open resource ↗plant-phenotyping.org · lines:307-379
Datasetpublic

The A data (RPi) were included as part of a larger citizen-powered study (“Leaf Targeting”, available at https://www.zooniverse.org/projects/venchen/leaf-targeting ) built on Zooniverse

Open resource ↗Zooniverse · venchen/leaf-targeting · lines:132-149

This is an automatically classified, unverified record. Curator approval is required before any resource enters the Catalog.