|
|
|
| Dynamic kernel perception for pavement distress detection in UAV inspection |
Jiangang ZHANG1( ),Xiao LI1,Dandan FENG2 |
1. School of Mathematics and Physics, Lanzhou Jiaotong University, Lanzhou 730070, China 2. School of Electronic and Information Engineering, Lanzhou Jiaotong University, Lanzhou 730070, China |
|
|
|
Abstract A detection model named dynamic kernel perception YOLO (DKP-YOLO) was constructed to address the problems of contextual information loss, multi-scale distress coexistence, and insufficient detail restoration in pavement distress detection from a UAV perspective. A lightweight crack context module was proposed to mitigate the variability of crack morphology and the loss of contextual information in complex scenes. This module enhanced feature discriminability through heterogeneous parallel depthwise separable convolutions and a dynamic feature fusion mechanism. In order to tackle the challenges of complex backgrounds and multi-scale distress, a pavement distress perception module was designed, which improved the multi-scale distress perception capability using parallel multi-scale convolutions and a dual-attention mechanism. A large-kernel perception feature fusion module was introduced to overcome the limitations of traditional convolutions in modeling long-range dependencies and responding to small targets. This module captured global contextual relationships by integrating extra-large receptive field convolution with channel-spatial attention. Finally, a lightweight dynamic upsampling operator was incorporated into the neck network to reduce detail loss and semantic ambiguity during upsampling. Experiments on the UAV-PDD2023 dataset showed that the proposed model significantly improved detection accuracy while effectively reducing model complexity. Compared with the baseline model, DKP-YOLO achieved a better balance between being lightweight and maintaining high detection performance, making it more suitable for UAV-based pavement distress detection.
|
|
Received: 09 December 2025
Published: 28 July 2026
|
|
|
| Fund: 甘肃省重点研发计划资助项目(25YFGA047);甘肃省基础研究创新群体资助项目(25JRRA805);鄂尔多斯市重点研发计划资助项目(YF20250245). |
基于动态核感知的无人机视角路面病害检测方法
针对无人机视角路面病害检测中上下文信息丢失、多尺度病害共存以及细节还原不足等问题,构建基于动态核感知的路面病害检测模型,DKP-YOLO. 为了缓解复杂场景裂缝形态多变与上下文信息易丢失的问题,提出轻量化裂缝上下文模块,该模块通过异构并行深度可分离卷积与动态特征融合机制增强特征区分度. 针对复杂背景与多尺度病害的挑战,设计路面病害感知模块,利用并行多尺度卷积与双重注意力机制提升多尺度病害感知能力. 为了克服传统卷积难以建模长程依赖和对微小目标响应不足的局限,提出大核感知特征融合模块,通过整合超大感受野卷积与通道-空间注意力捕捉全局上下文关系. 为了减少上采样过程中细节丢失与语义模糊,在颈部网络引入轻量级动态上采样算子. 在UAV-PDD2023数据集上的实验表明,本研究模型在显著提升检测精度的同时,实现了模型复杂度的有效降低. 与基线模型相比,本研究模型在轻量化和检测性能上取得更好的平衡,更适合无人机视角路面病害目标检测.
关键词:
路面病害检测,
无人机航拍图像,
动态核感知,
多尺度特征融合,
复杂背景抑制
|
|
| [1] |
中华人民共和国交通运输部. 2024年交通运输行业发展统计公报[EB/OL]. (2025–06–12) [2025–07–09]. https://xxgk.mot.gov.cn/2020/jigou/zhghs/202506/t20250610_4170228.html.
|
|
|
| [2] |
《中国公路学报》编辑部 中国路面工程学术研究综述·2024[J]. 中国公路学报, 2024, 37 (3): 1- 49 Editorial Department of China Journal of Highway and Transport Review on China’s pavement engineering research: 2024[J]. China Journal of Highway and Transport, 2024, 37 (3): 1- 49
doi: 10.19815/j.jace.2025.04061
|
|
|
| [3] |
ALZAHRANI B, OUBBATI O S, BARNAWI A, et al UAV assistance paradigm: state-of-the-art in applications and challenges[J]. Journal of Network and Computer Applications, 2020, 166: 102706
doi: 10.1016/j.jnca.2020.102706
|
|
|
| [4] |
ZHANG Y, ZUO Z, XU X, et al Road damage detection using UAV images based on multi-level attention mechanism[J]. Automation in Construction, 2022, 144: 104613
doi: 10.1016/j.autcon.2022.104613
|
|
|
| [5] |
WANG F, ZOU Y, CHEN X, et al Rapid in-flight image quality check for UAV-enabled bridge inspection[J]. ISPRS Journal of Photogrammetry and Remote Sensing, 2024, 212: 230- 250
doi: 10.1016/j.isprsjprs.2024.05.008
|
|
|
| [6] |
MA X, LI Y, YANG Z, et al Lightweight network for millimeter-level concrete crack detection with dense feature connection and dual attention[J]. Journal of Building Engineering, 2024, 94: 109821
doi: 10.1016/j.jobe.2024.109821
|
|
|
| [7] |
杨豪, 刘李彦, 张军辉, 等 环境适应性优化的轻量化多尺度道路裂缝检测[J]. 中国公路学报, 2025, 38 (7): 118- 134 YANG Hao, LIU Liyan, ZHANG Junhui, et al Environmental adaptability optimization lightweight multi-scale road crack detection[J]. China Journal of Highway and Transport, 2025, 38 (7): 118- 134
|
|
|
| [8] |
MA X, DAI X, BAI Y, et al. Rewrite the stars [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Seattle: IEEE, 2024: 5694–5703.
|
|
|
| [9] |
YANG F, ZHANG L, YU S, et al Feature pyramid and hierarchical boosting network for pavement crack detection[J]. IEEE Transactions on Intelligent Transportation Systems, 2020, 21 (4): 1525- 1535
doi: 10.1109/TITS.2019.2910595
|
|
|
| [10] |
LIU Z, LIN Y, CAO Y, et al. Swin transformer: hierarchical vision transformer using shifted windows [C]// Proceedings of the IEEE/CVF International Conference on Computer Vision. Montreal: IEEE, 2022: 9992–10002.
|
|
|
| [11] |
WANG L, YOON K J Knowledge distillation and student-teacher learning for visual intelligence: a review and new outlooks[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2022, 44 (6): 3048- 3068
doi: 10.1109/TPAMI.2021.3055564
|
|
|
| [12] |
LIU H, MIAO X, MERTZ C, et al. CrackFormer: transformer network for fine-grained crack detection [C]// Proceedings of the IEEE/CVF International Conference on Computer Vision. Montreal: IEEE, 2022: 3763–3772.
|
|
|
| [13] |
DUNG C V, ANH L D Autonomous concrete crack detection using deep fully convolutional neural network[J]. Automation in Construction, 2019, 99: 52- 58
doi: 10.1016/j.autcon.2018.11.028
|
|
|
| [14] |
QU Z, CHEN W, WANG S Y, et al A crack detection algorithm for concrete pavement based on attention mechanism and multi-features fusion[J]. IEEE Transactions on Intelligent Transportation Systems, 2022, 23 (8): 11710- 11719
doi: 10.1109/TITS.2021.3106647
|
|
|
| [15] |
PAN Y, ZHANG G, ZHANG L A spatial-channel hierarchical deep learning network for pixel-level automated crack detection[J]. Automation in Construction, 2020, 119: 103357
doi: 10.1016/j.autcon.2020.103357
|
|
|
| [16] |
TAN M, PANG R, LE Q V. EfficientDet: scalable and efficient object detection [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Seattle: IEEE, 2020: 10778–10787.
|
|
|
| [17] |
HOU Q, ZHOU D, FENG J. Coordinate attention for efficient mobile network design [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville: IEEE, 2021: 13708–13717.
|
|
|
| [18] |
SHAN J, JIANG W, FENG X Bridging cross-domain and cross-resolution gaps for UAV-based pavement crack segmentation[J]. Automation in Construction, 2025, 174: 106141
doi: 10.1016/j.autcon.2025.106141
|
|
|
| [19] |
TIAN Y, YE Q, DOERMANN D. YOLOv12: attention-centric real-time object detectors [EB/OL]. [2025–02–18]. https://arxiv.org/abs/2502.12524.
|
|
|
| [20] |
ZHAO H, KONG X, HE J, et al. Efficient image super-resolution using pixel attention [C]// Computer Vision – ECCV 2020 Workshops. Cham: Springer, 2020: 56–72.
|
|
|
| [21] |
LAU K W, PO L M, REHMAN Y A U Large separable kernel attention: rethinking the large kernel attention design in CNN[J]. Expert Systems with Applications, 2024, 236: 121352
doi: 10.1016/j.eswa.2023.121352
|
|
|
| [22] |
LIU W, LU H, FU H, et al. Learning to upsample by learning to sample [C]// Proceedings of the IEEE/CVF International Conference on Computer Vision. Paris: IEEE, 2024: 6004–6014.
|
|
|
| [23] |
YAN H, ZHANG J UAV-PDD2023: a benchmark dataset for pavement distress detection based on UAV images[J]. Data in Brief, 2023, 51: 109692
doi: 10.1016/j.dib.2023.109692
|
|
|
| [24] |
ZHAO Y, LV W, XU S, et al. DETRs beat YOLOs on real-time object detection [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Seattle: IEEE, 2024: 16965–16974.
|
|
|
| [25] |
WANG C Y, YEH I H, MARK LIAO H Y. YOLOv9: learning what you want toLearn using programmable gradient information [C]//Computer Vision – ECCV 2024. Cham: Springer, 2025: 1–21.
|
|
|
| [26] |
WANG A, CHEN H, LIU L H, et al. YOLOv10: real-time end-to-end object detection [C]// 38th Conference on Neural Information Processing Systems. Vancouver: Curran Associates, 2024: 107984–108011.
|
|
|
| [27] |
ZHU J, ZHONG J, MA T, et al Pavement distress detection using convolutional neural networks with images captured via UAV[J]. Automation in Construction, 2022, 133: 103991
doi: 10.1016/j.autcon.2021.103991
|
|
|
|
Viewed |
|
|
|
Full text
|
|
|
|
|
Abstract
|
|
|
|
|
Cited |
|
|
|
|
| |
Shared |
|
|
|
|
| |
Discussed |
|
|
|
|