An improved SSD method for infrared target detection based on convolutional neural network

Author:

Liu Gang,Cao Zixuan,Liu Sen,Song Bin,Liu Zhonghua

Abstract

Target detection is the basis for automatic target recognition system of infrared imaging guidance to complete subsequent tasks such as recognition and tracking. Existing systems have not the autonomous learning ability of target feature, and it will be powerless once the task environment exceeds the pre-planned condition. The single-stage target detection based on deep learning has the ability of autonomous learning and high computational efficiency, which is an effective way to solve the problem of infrared imaging guidance target detection in complex environment. SSD (Single Shot MultiBox Detector) is a classical single-stage detection model, however, the convolution layer with strong semantic information in SSD has low resolution, which is not conducive to small target detection. In addition, the location loss of SSD does not consider the impact of target scale change. Therefore, this paper puts forward two improvement ideas in view of SSD: (1) Starting from FPN (Feature Pyramid Network), feature channel’s importance is distinguished through efficient channel attention mechanism, the contribution of each feature layer to the fusion output is described based on the learnable weight, and the feature weighted fusion of bidirectional multi-scale is realized between the feature layer which has low resolution and strong semantics and the feature layer which has high resolution and weak semantics. (2) Starting from IoU (Intersection over Union) and considering non overlapping parts and geometric relationship between the predicted box and the ground-truth box, the location loss of SSD that remains invariable to the target scale change is constructed to improve the sensitivity of the detection model to the locating error of small target. The experimental results show that, for 300 × 300 input, the presented method achieves 84.7% mAP (mean Average Precision) on VOC2007 test and for 512 × 512 input, it reaches 86.6%. On the self-built infrared aircraft data set, the proposed method achieves 81.1% mAP and can detect more small targets. Without affecting detection speed, the presented method on experimental results outperforms some comparable state-of-the-art models such as YOLOv3 (You Only Look Once), DSSD (Deconvolutional Single Shot Multibox Detector), RSSD (Rainbow Single Shot Multibox Detector) and FSSD (Fusion Single Shot Multibox Detector).

Publisher

IOS Press

Subject

Computational Mathematics,Computer Science Applications,General Engineering

Reference31 articles.

1. The role of the primary visual cortex in higher level vision;Sing;Vision Research,1998

2. Deep machine learning-a new frontier in artificial intelligence research;Arel;IEEE Transactions on Computational Intelligence Magazine,2010

3. Deep learning review;LeCun;Nature,2015

4. Advances and perspectives on applications of deep learning in visual object detection;Zhang;Acta Automatica Sinica,2017

5. Deep learning for generic object detection: a survey;Liu;International Journal of Computer Vision,2020

Cited by 1 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

1. Research Progress of Water Surface Object Detection;2023 IEEE/ACIS 23rd International Conference on Computer and Information Science (ICIS);2023-06-23

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3