Aug 2026· Electronics· Vol 15, pp. 3630· 0 citations
TL;DR
Compared with state-of-the-art lightweight detectors including YOLOv11n and Hyper-YOLO, YOLO-STCDE attains the highest detection accuracy with the smallest model footprint, demonstrating an optimal balance between accuracy and efficiency for real-world bridge crack inspection.
Abstract
Bridge crack detection is a critical task in structural health monitoring, yet existing deep learning methods often suffer from parameter redundancy in backbone networks, insufficient adaptivity to complex background interference in feature enhancement modules, and feature conflicts between classification and regression subtasks in coupled detection heads. To address these challenges, this paper proposes YOLO-STCDE, an improved lightweight detection model built upon YOLOv12. The model integrates three synergistic architectural innovations: (1) ST-Net, a lightweight backbone that replaces the original R-ELAN with a star operation-based design incorporating DynamicTanh activation for implicit high-dimensional feature mapping and adaptive amplitude calibration; (2) A2C2f-CD, a dual-dynamic gated neck module that embeds DynamicTanh and Convolutional Gated Linear Units into the A2C2f architecture, enhancing crack feature discrimination under complex backgrounds; and (3) Efficient-Detect, a decoupled detection head with a shared convolution stem that eliminates task conflict while substantially compressing parameter overhead. Extensive experiments on the bridge crack dataset demonstrate that YOLO-STCDE achieves 91.6% mAP@0.5 and 69.4% mAP@0.5:0.95 with only 2.25 M parameters and 5.6 GFLOPs, representing improvements of 3.2 and 9.5 percentage points over the YOLOv12n baseline, respectively, while simultaneously reducing the parameter count by 10.4%. Compared with state-of-the-art lightweight detectors including YOLOv11n and Hyper-YOLO, YOLO-STCDE attains the highest detection accuracy with the smallest model footprint, demonstrating an optimal balance between accuracy and efficiency for real-world bridge crack inspection.
A multi-module collaborative lightweight model (MCL-YOLO) based on YOLOv12 is proposed to reduce computational complexity while preserving critical information during feature downsampling to enhance the representation of slender, curved, and branched crack patterns.
Bing-Yu Han, Yang Wu, Wen-Hao Feng et al.· Italian National Conference...· 0 citations
Surface cracks are critical indicators of deterioration in flood-control infrastructure, yet automated detection from inspection imagery remains challenging due to complex backgrounds, elongated geometries, and variations in apparent scale. This study aims to develop a lightweight detector for accurate dike crack detec...
Automated bridge crack detection is challenging because cracks often exhibit weak contrast, irregular morphology, slender structures, and strong interference from complex surface textures. To address these issues, this study proposes a YOLOv8-CA-EYHL framework that combines filtering-equalization preprocessing with coo...
Xian-Wei Zhu, He Chao, Ya-Hui Zhang· PLoS ONE· 0 citations
To address the limitations of YOLOv8n in steel surface defect detection–namely, its large parameter count and computational cost, insufficient multi-scale feature modeling capability, and poor scale adaptability of the detection head– we propose a lightweight improved model named EDGI-YOLO. The model is restructured al...
Pavement surface distress detection is an important task in road maintenance and intelligent infrastructure inspection. In practical vehicle-mounted inspection images, cracks and other distress targets often present weak edges, irregular shapes, large scale variations, and strong background interference, which makes st...
Peng Li, Tianyang Wang, Lu-Sheng Liu et al.· International Conference on...· 0 citations
This network alleviates the trade-off between detection accuracy and computational efficiency, enabling efficient and robust detection of multi-scale defects on metal substrates and substantially reduces parameter redundancy while maintaining multi-scale feature representation capabilities.