Aug 2026· Measurement science and technology· Vol 37· 0 citations· 68 references
Physics
TL;DR
Compared to existing object detection models, SOCD-YOLO demonstrates enhanced performance in terms of citrus fruit detection accuracy and robustness, providing a valuable reference for artificial intelligence based real-time fruit detection and position measurement in densely occluded environments.
Abstract
To address the challenges of low detection accuracy, high miss rates, and limited model lightweightness arising from dense fruit distribution, foliage occlusion, and small fruit size during citrus fruit localisation and recognition in complex orchard environments, this study proposes a lightweight small-object citrus fruit detection model based on an improved YOLO11 architecture, termed SOCD-YOLO (YOLO for small object citrus detection). Firstly, a DLTBlock is designed by integrating element-wise multiplication with a triplet attention mechanism to reconstruct the C3k2 module, thereby enhancing the nonlinear fusion of high-dimensional features. This design effectively suppresses interference from occluding foliage and complex backgrounds, and improving the robustness of fruit target recognition. Secondly, the traditional downsampling operation is replaced with the adaptive downsampling block module, which effectively alleviates information loss during feature propagation for small-object features, while simultaneously reducing model parameters and computing complexity, thus enhancing small-object recognition accuracy. Finally, a lightweight P2-focused pyramid structure is constructed to further enhance the model’s capability in identifying small and densely distributed objects, while significantly reducing the missed detection rate. Experimental findings indicate that, on the CitDet dataset, the proposed model achieves improvements of 4.3%, 6.3%, and 5.6% in Precision, Recall, and mean average precision, respectively, while reducing the parameter count and model size to 1.5 M and 3.5 MB. Moreover, the FPS reaches 103.5 frames/s. On the Tomato and PASCAL VOC 2007 datasets, overall performance is consistently improved. Compared to existing object detection models, SOCD-YOLO demonstrates enhanced performance in terms of citrus fruit detection accuracy and robustness, providing a valuable reference for artificial intelligence based real-time fruit detection and position measurement in densely occluded environments.
In unstructured orchard environments, detecting immature green citrus faces challenges such as background interference of similar colors, severe occlusion by branches and leaves, and dense distributions of small targets. These factors lead to low detection accuracy and poor performance in real time. To address these is...
Jun-Tao Xiong, Kun Tang, Chang Lu et al.· International Conference on...· 0 citations
Manual fruit thinning is labor-intensive and inefficient, making the development of intelligent visual detection systems a crucial approach for improving the quality and production efficiency of the pear industry. However, in natural orchard environments, young pear fruits are small in size with slender fruit stalks, a...
Tian-Zhao Jian, Xiu-Hua Zhang, De-Hai Kong et al.· Agriculture· 0 citations
In the process of agricultural intelligence, precise detection of plant organs serves as the foundation for core tasks such as crop phenotyping analysis and yield prediction. However, in complex field environments, small targets such as citrus flowers and shoots face challenges including scale variation, background int...
Wen-Feng Guo, Zhi-Fang Bi, Lin-Juan Wang et al.· INMATEH Agricultural Enginee...· 0 citations
Reliable joint detection of mango fruits and stems is an essential upstream perception task for robotic harvesting, but remains challenging because stems are small, slender, frequently occluded, and visually degraded by illumination variation. This study proposes MangoNET, a YOLOv11n-based framework for joint mango fru...
This work proposes Enhance-YOLOv8, which replaces YOLOv8's Cross Stage Partial with 2 convolutions (C2f) backbone module with Enhance Adaptive Fine-grained Channel Attention (Enhance_AFCA), which integrates hierarchical multi-scale feature extraction and adaptive edge enhancement to address edge information loss and in...
Saiqi Pi, Fa-Yuan Xu, Fei Wang et al.· PeerJ Computer Science· 0 citations
To address the limitations of computational resources and the stringent real-time stability requirements in post-harvest online grading of fresh tea leaves, this study proposes a lightweight object detection method oriented toward efficient deployment. The YOLOv11n network was selected as the baseline model. A StarNet...