Unverified paper record
An improved YOLOv8-seg-based method for key part segmentation of tobacco plants
Frontiers in plant science · 22 Sept 2025 · 10.3389/fpls.2025.1673202
Abstract
Accurate segmentation of key tobacco structures is essential for enabling automated harvesting. However, complex backgrounds, variable lighting conditions, and blurred boundaries between the stem and petiole significantly hinder segmentation accuracy in field environments. To overcome these challenges, we propose an enhanced instance segmentation approach based on YOLOv8-seg, incorporating depth-based background filtering and architectural improvements. Specifically, depth information from RGB-D images is employed to spatially filter non-target background regions, thereby enhancing foreground clarity. In addition, a Hybrid Dilated Residual Attention Block (HDRAB) is integrated into the YOLOv8-seg backbone to improve boundary discrimination between petioles and stems, while a Lightweight Shared Detail-Enhanced Convolution Detection Head (LSDECD) is designed to efficiently capture fine-grained texture features. Experimental results demonstrate that depth filtering increases mAP50 bb and mAP50 seg by 7.9% and 6.3%, respectively, while the architectural enhancements further raise them to 89.5% and 91.1%, surpassing the YOLOv8-seg baseline by 5.2% and 10.0%. Compared with mainstream models such as Mask R-CNN and SOLOv2, the proposed method achieves superior segmentation accuracy with low computational cost, highlighting its potential for practical deployment in automated tobacco harvesting.
Plant phenotyping relevance
タバコ植物の茎・葉柄などの構造をRGB-D画像からセグメンテーションする手法を開発・比較しており、植物器官形態の取得が中心である。
abstractwe propose an enhanced instance segmentation approach based on YOLOv8-seg, incorporating depth-based background filtering and architectural improvements.
abstractdepth information from RGB-D images is employed to spatially filter non-target background regions
abstractimprove boundary discrimination between petioles and stems
Code and data availability
The supplied blocks describe a custom tobacco RGB-D image dataset (2,366 images, Labelme annotations) and an improved YOLOv8-seg model, but contain no data availability statement, code deposit, or authors' public URL for the dataset, annotations, or trained model. The only URL present is the Intel RealSense D435 camera
No evidence-backed public reproduction asset is currently recorded.
This is an automatically classified, unverified record. Curator approval is required before any resource enters the Catalog.