Selective Cotton Boll Localization for Robotic Harvesting: Evaluation of Deep Learning Vision Models Under Field Conditions

📅 2026-09-16
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
研究开发并评估了基于深度学习的棉花选择性采摘视觉框架,使用YOLO系列模型解决自然条件下棉花识别与定位问题。
📝 Abstract
This study developed and evaluated a deep-learning-based perception framework for selective robotic cotton picking. The dataset contained 1,008 annotated field images collected using three cameras under varying natural lighting and weather conditions. Object-detection models from the YOLOv8 through YOLOv13 families were evaluated using their default configurations, while segmentation performance was assessed using YOLOv8-seg, YOLOv11-seg, YOLOv12-seg, the Segment Anything Model (SAM), SAMv2.1, FastSAM, and Grounded-SAM with the Recognize Anything Model (RAM). Among the detection models, GELAN-s achieved the most favorable balance between mean average precision (mAP) and inference speed, obtaining an mAP of 86.1%, precision of 81.6%, recall of 76.6%, and an F1-score of 79.0%, with an average inference time of 42.3 ms per image. Among the direct segmentation models, YOLOv12-m-seg provided the most favorable balance between AP@0.5 and FPS, achieving a segmentation AP@0.5 of 83.7% with an inference time of 20.4 ms per image. In the detection-prompted segmentation approach, bounding-box prompts generated by GELAN-s improved the localization of cotton bolls for SAM and SAMv2.1, while SAMv2.1 Tiny consistently outperformed FastSAM and Grounded-SAM with RAM. In the area-based evaluation against manually annotated segmentation masks, YOLOv12-m-seg achieved an $R^2$ value of 0.966, compared with 0.860 for GELAN-s + SAMv2.1 Tiny. Field experiments conducted using a UR5e robotic manipulator, a custom end-effector, and a ZED2i stereo camera further validated the effectiveness of the YOLOv12-m-seg model for real-time cotton boll detection, segmentation, and selective picking under varying confidence levels. These results demonstrate that YOLOv12-m-seg provides an efficient perception model for robotic cotton harvesting and has strong potential for field deployment.
Problem

Research questions and friction points this paper is trying to address.

Selective Cotton Boll Localization
Robotic Harvesting
Deep Learning Vision Models
Field Conditions
Innovation

Methods, ideas, or system contributions that make the work stand out.

Deep Learning
Selective Robotic Cotton Picking
YOLOv12-m-seg
GELAN-s
Segmentation
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
T
Thevathayarajh Thayananthan
School of Environmental, Civil, Agricultural and Mechanical Engineering, University of Georgia, Athens, 30602, GA, USA
X
Xin Zhang
School of Environmental, Civil, Agricultural and Mechanical Engineering, University of Georgia, Athens, 30602, GA, USA
I
Isuru Laddusinghe Badu
School of Environmental, Civil, Agricultural and Mechanical Engineering, University of Georgia, Athens, 30602, GA, USA
J
Jonathan Harjono
School of Environmental, Civil, Agricultural and Mechanical Engineering, University of Georgia, Athens, 30602, GA, USA
G
Glen C. Rains
Department of Entomology, University of Georgia, Tifton, 31793, GA, USA
Beiwen Li
Beiwen Li
Associate Professor of Mechanical Engineering, University of Georgia
3D optical metrologysuperfast 3D imagingin-situ inspectionfringe analysis3D imaging
L
Leonardo M. Bastos
Department of Crop and Soil Sciences, University of Georgia, Athens, 30602, GA, USA
N
Nuwan K. Wijewardane
Department of Agricultural and Biological Engineering, Mississippi State University, Mississippi State, 39762, MS, USA
V
Vitor S. Martins
Department of Agricultural and Biological Engineering, Mississippi State University, Mississippi State, 39762, MS, USA