PGDS-YOLOv8s: An Improved YOLOv8s Model for Object Detection in Fisheye Images-Reference-Cited by-同舟云学术

PGDS-YOLOv8s: An Improved YOLOv8s Model for Object Detection in Fisheye Images

Published:2023-12-20 Issue:1 Volume:14 Page:44
ISSN:2076-3417
Container-title:Applied Sciences
language:en
Short-container-title:Applied Sciences

Author:

Yang Degang¹^ORCID,Zhou Jie¹,Song Tingting¹,Zhang Xin¹,Song Yingze¹

Affiliation:

1. College of Computer and Information Science, Chongqing Normal University, Chongqing 401331, China

Abstract

Recently, object detection has become a research hotspot in computer vision, which often detects regular images with small viewing angles. In order to obtain a field of view without blind spots, fisheye cameras, which have distortions and discontinuities, have come into use. The fisheye camera, which has a wide viewing angle, and an unmanned aerial vehicle equipped with a fisheye camera are used to obtain a field of view without blind spots. However, distorted and discontinuous objects appear in the captured fisheye images due to the unique viewing angle of fisheye cameras. It poses a significant challenge to some existing object detectors. To solve this problem, this paper proposes a PGDS-YOLOv8s model to solve the issue of detecting distorted and discontinuous objects in fisheye images. First, two novel downsampling modules are proposed. Among them, the Max Pooling and Ghost’s Downsampling (MPGD) module effectively extracts the essential feature information of distorted and discontinuous objects. The Average Pooling and Ghost’s Downsampling (APGD) module acquires rich global features and reduces the feature loss of distorted and discontinuous objects. In addition, the proposed C2fs module uses Squeeze-and-Excitation (SE) blocks to model the interdependence of the channels to acquire richer gradient flow information about the features. The C2fs module provides a better understanding of the contextual information in fisheye images. Subsequently, an SE block is added after the Spatial Pyramid Pooling Fast (SPPF), thus improving the model’s ability to capture features of distorted, discontinuous objects. Moreover, the UAV-360 dataset is created for object detection in fisheye images. Finally, experiments show that the proposed PGDS-YOLOv8s model on the VOC-360 dataset improves mAP@0.5 by 19.8% and mAP@0.5:0.95 by 27.5% compared to the original YOLOv8s model. In addition, the improved model on the UAV-360 dataset achieves 89.0% for mAP@0.5 and 60.5% for mAP@0.5:0.95. Furthermore, on the MS-COCO 2017 dataset, the PGDS-YOLOv8s model improved AP by 1.4%, AP50 by 1.7%, and AP75 by 1.2% compared with the original YOLOv8s model.

Funder

Natural Science Foundation of Chongqing

Science and Technology Research Program of Chongqing Municipal Education Commission

Chongqing Normal University Ph.D. Start-up Fund

Publisher

MDPI AG

Subject

Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science

Link

https://www.mdpi.com/2076-3417/14/1/44/pdf

Reference46 articles.

1. Song, J., Yu, Z., Qi, G., Su, Q., Xie, J., and Liu, W. (2023). UAV Image Small Object Detection Based on RSAD Algorithm. Appl. Sci., 13.

2. Mou, C., Liu, T., Zhu, C., and Cui, X. (2023). WAID: A Large-Scale Dataset for Wildlife Detection with Drones. Appl. Sci., 13.

3. Barmpoutis, P., Stathaki, T., Dimitropoulos, K., and Grammalidis, N. (2020). Early fire detection based on aerial 360-degree sensors, deep convolution neural networks and exploitation of fire dynamic textures. Remote Sens., 12.

4. Autonomous detection of damage to multiple steel surfaces from 360 panoramas using deep neural networks;Luo;Comput. Aided Civ. Infrastruct. Eng.,2021

5. Autonomous aerial robot using dual-fisheye cameras;Gao;J. Field Robot.,2020

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Early warning system for nocardiosis in largemouth bass (Micropterus salmoides) based on multimodal information fusion;Computers and Electronics in Agriculture;2024-11

2. A Deep-Learning-Based Model for the Detection of Diseased Tomato Leaves;Agronomy;2024-07-22

3. Improved YOLOv8 Model for a Comprehensive Approach to Object Detection and Distance Estimation;IEEE Access;2024

4. Early Warning System for Nocardiosis in Largemouth Bass (Micropterus Salmoides) Based on Multimodal Information Fusion;2024