A Block-Based and Highly Parallel CNN Accelerator for Seed Sorting-Reference-Cited by-同舟云学术

A Block-Based and Highly Parallel CNN Accelerator for Seed Sorting

Published:2022-11-17 Issue: Volume:2022 Page:1-16
ISSN:2090-0155
Container-title:Journal of Electrical and Computer Engineering
language:en
Short-container-title:Journal of Electrical and Computer Engineering

Author:

Sang Xiaoting¹^ORCID,Hu Zhenghui²^ORCID,Li Huanyu³^ORCID,Li Chunlei¹^ORCID,Liu Zhoufeng¹^ORCID

Affiliation:

1. School of Electronic and Information Engineering, Zhongyuan University of Technology, Zhengzhou 450007, China

2. Hangzhou Innovation Institute, Beihang University, Hangzhou 310051, China

3. College of Oceanography and Space Informatics, China University of Petroleum East China, Qingdao 266580, China

Abstract

Seed sorting is critical for the breeding industry to improve the agricultural yield. The seed sorting methods based on convolutional neural networks (CNNs) have achieved excellent recognition accuracy on large-scale pretrained network models. However, CNN inference is a computationally intensive process that often requires hardware acceleration to operate in real time. For embedded devices, the high-power consumption of graphics processing units (GPUs) is generally prohibitive, and the field programmable gate array (FPGA) becomes a solution to perform high-speed inference by providing a customized accelerator for a particular user. To date, the recognition speeds of the FPGA-based universal accelerators for high-throughput seed sorting tasks are slow, which cannot guarantee real-time seed sorting. Therefore, a block-based and highly parallel MobileNetV2 accelerator is proposed in this paper. First, a hardware-friendly quantization method that uses only fixed-point operation is designed to reduce resource consumption. Then, the block convolution strategy is proposed to avoid latency and energy consumption increase caused by large-scale intermediate result off-chip data transfers. Finally, two scalable computing engines are explicitly designed for depth-wise convolution (DWC) and point-wise convolution (PWC) to develop the high parallelism of block convolution computation. Moreover, an efficient memory system with a double buffering mechanism and new data reordering mode is designed to address the imbalance between memory access and parallel computing. Our proposed FPGA-based MobileNetV2 accelerator for real-time seed sorting is implemented and evaluated on the platform of Xilinx XC7020. Experimental results demonstrate that our implementation can achieve about 29.4 frames per second (FPS) and 10.86 Giga operations per second (GOPS), and 0.92× to 5.70 × DSP-efficiency compared with previous FPGA-based accelerators.

Funder

National Natural Science Foundation of China

Publisher

Hindawi Limited

Subject

Electrical and Electronic Engineering,General Computer Science,Signal Processing

Link

http://downloads.hindawi.com/journals/jece/2022/5608573.pdf

Reference63 articles.

1. Discriminating and elimination of damaged soybean seeds based on image characteristics

2. Classification of haploid and diploid maize seeds by using image processing techniques and support vector machines;Y. Altuntaş

3. Support vector machines

4. Linear discriminant analysis;A. J. Izenman,2013

5. Artificial neural network;S. C. Wang,2003

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. EfficientMaize: A Lightweight Dataset for Maize Classification on Resource-Constrained Devices;Data in Brief;2024-06

2. Real‐time seed sorting system via 2D information entropy‐based CNN pruning and TensorRt acceleration;IET Image Processing;2023-01-30