missing detection and low detection accuracy in single shot multibox detector(SSD)
an improved SSD object detection algorithm is proposed. The conv4_3 convolution layer and the previous two standard convolutions are replaced by dilated convolution to expand the receptive field. The deconvolution is used to fuse the feature maps of different scales
so that the feature maps formed by fusion contain rich context information. The attention model is added to the feature map to effectively extract the features of regions of interest. Simulation results show that the detection accuracy of the improved algorithm is 0.9% higher than that of the original algorithm
and the detection effect is better. To some extent
it solves the problems of false detection and missing detection
WANG Haoxian, DONG Heng, ZHOU Zhiquan. Review on dim small target detection technologies in infrared single frame images [J]. Laser & Optoelectronics Progress, 2019, 56(8): 1-14.
OU Pan, ZHANG Zheng, LU Kui, et al. Object detection of remote sensing images based on convolutional neural networks [J]. Laser Optoelectronics Progress, 2019, 56(5): 051002.
YANG Jie, CHEN Lingna, CHEN Yushao, et al. Target detection and recognition based on depth learning [J]. Information Technology, 2018, 42(10): 81-87.
LOWE D G. Distinctive image features from scale-invariant keypoints [J]. International Journal of Computer Vision, 2004, 60(2): 91-110.
VIOLA P, JONES M J. Robust real-time face detection [J]. International Journal of Computer Vision, 2004, 57(2): 137-154.
GIRSHICK R, DONAHUE J, DARRELL T, et al. Rich feature hierarchies for accurate object detection and semantic segmentation [C]∥Proceedings of 2014 IEEE Conference on Computer Vision and Pattern Recognition. Piscataway, NJ, USA: IEEE, 2014: 580-587.
GIRSHICK R. Fast R-CNN [C]∥Proceedings of 2015 IEEE International Conference on Computer Vision. Piscataway, NJ, USA: IEEE, 2015: 1440-1448.
REN Shaoqing, HE Kaiming, GIRSHICK R, et al. Faster R-CNN: towards real-time object detection with region proposal networks [J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2017, 39(6): 1137-1149.
LIN C F, WANG S D. Fuzzy support vector machines [J]. IEEE Transactions on Neural Networks, 2002, 13(2): 464-471.
REDMON J, DIVVALA S, GIRSHICK R, et al. You only look once: unified, real-time object detection [C]∥Proceedings of 2016 IEEE Conference on Computer Vision and Pattern Recognition. Piscataway, NJ, USA: IEEE, 2016: 779-788.
LIU W, ANGUELOV D, ERHAN D, et al. SSD: single shot MultiBox detector [C]∥Proceedings of the 14th European Conference on Computer Vision. Berlin, Germany: Springer, 2016: 21-37.
SIMONYAN K, ZISSERMAN A. Very deep convolutional networks for large-scale image recognition [EB/OL].[2020-09-26]. http:∥arXiv.org/abs/1409.155 6v6.
GIRSHICK R, DONAHUE J, DARRELL T, et al. Rich feature hierarchies for accurate object detection and semantic segmentation [C]∥IEEE Conference on Computer Vision and Pattern Recognition. Piscataway, NJ, USA: IEEE Computer Society, 2014: 580-587.
YU F, KOLTUN V. Multi-scale context aggregation by dilated convolutions [EB/OL].[2020-10-01]. https:∥arXiv.org/abs/1511.07122.
PENG L Y, ZHANG L, ZHANG Y, et al. Deep deconvolution neural network for image super-resolution [J]. Journal of Software, 2018, 29(4): 927-934.
JADERBERG M, SIMONYAN K, ZISSERMAN A, et al. Spatial transformer networks [C]∥Advances in Neural Information Processing Systems. Montreal, Canada: [s.n.], 2015: 2008-2016.
HU Jie, SHEN Li, SUN Gang. Squeeze-and-excitation networks [C]∥Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. Piscataway, NJ, USA: IEEE, 2018: 7132-7141.
WANG Fei, JIANG Mengqing. Residual attention network for image classification [EB/OL]. [2020-09-22]. https:∥arXiv.org/abs/1704.06904
FU Chengyang, LIU Wei, RANGA A, et al. DSSD: deconvolutional single shot detector [EB/OL]. [2020-09-22]https:∥arXiv.org/abs/1701.06659.