MSD-YOLO11n: an improved small target detection model for high precision UAV aerial imagery
摘要
The capacity for precise target detection in UAV aerial images represents a pivotal prerequisite for the advancement of low-altitude economic activities. The technology for large target detection in images has attained a state of relative maturity, while the identification of small targets remains encumbered by challenges such as indistinct edge information, the absence of comprehensive deep feature information, and the constrained capacity for the expression of detection head information. To address these challenges, the MSD-YOLO11n target detection model has been proposed. Firstly, the P2 layer downsampling convolution is adopted as SPDConv, while the Multiscale Edge Information Selection (MEIS) module is proposed to replace the residual module in C3K2 in order to enhance the edge information feature extraction of small targets. Secondly, the Smalltarget Feature Enhancement Pyramid (SFEP) module has been proposed as a means of achieving the fusion of feature information between the P2 and P3 layers. This is achieved by combining Dysample up sampling with SPDConv down sampling and reconstructing low-quality images to obtain high-quality images using the CSP-OmniKernel module. Subsequently, the DyHead module was utilised to integrate scale, space and task-aware attention. Finally, the proposed MSD-YOLO11n model was evaluated based on the VisDrone2019 dataset through a combination of ablation and comparison experiments. In comparison with YOLO11n, both