Element detection and segmentation of mathematical function graphs based on improved Mask R-CNN-Reference-Cited by-同舟云学术

Element detection and segmentation of mathematical function graphs based on improved Mask R-CNN

Published:2023 Issue:7 Volume:20 Page:12772-12801
ISSN:1551-0018
Container-title:Mathematical Biosciences and Engineering
language:
Short-container-title:MBE

Author:

Lu Jiale,Chen Jianjun,Xu Taihua,Song Jingjing,Yang Xibei

Abstract

<abstract> <p>There are approximately 2.2 billion people around the world with varying degrees of visual impairments. Among them, individuals with severe visual impairments predominantly rely on hearing and touch to gather external information. At present, there are limited reading materials for the visually impaired, mostly in the form of audio or text, which cannot satisfy the needs for the visually impaired to comprehend graphical content. Although many scholars have devoted their efforts to investigating methods for converting visual images into tactile graphics, tactile graphic translation fails to meet the reading needs of visually impaired individuals due to image type diversity and limitations in image recognition technology. The primary goal of this paper is to enable the visually impaired to gain a greater understanding of the natural sciences by transforming images of mathematical functions into an electronic format for the production of tactile graphics. In an effort to enhance the accuracy and efficiency of graph element recognition and segmentation of function graphs, this paper proposes an MA Mask R-CNN model which utilizes MA ConvNeXt as its improved feature extraction backbone network and MA BiFPN as its improved feature fusion network. The MA ConvNeXt is a novel feature extraction network proposed in this paper, while the MA BiFPN is a novel feature fusion network introduced in this paper. This model combines the information of local relations, global relations and different channels to form an attention mechanism that is able to establish multiple connections, thus increasing the detection capability of the original Mask R-CNN model on slender and multi-type targets by combining a variety of multi-scale features. Finally, the experimental results show that MA Mask R-CNN attains an 89.6% mAP value for target detection and 72.3% mAP value for target segmentation in the instance segmentation of function graphs. This results in a 9% mAP improvement for target detection and 12.8% mAP improvement for target segmentation compared to the original Mask R-CNN.</p> </abstract>

Publisher

American Institute of Mathematical Sciences (AIMS)

Subject

Applied Mathematics,Computational Mathematics,General Agricultural and Biological Sciences,Modeling and Simulation,General Medicine

Reference48 articles.

1. World Health Organization, World Report on Vision, 2019. Available from: https://www.who.int/publications/i/item/world-report-on-vision.

2. K. Han, Y. Liu, F. Yuan, The tactile graphics representation and its design application for the visually impaired, Design, 4 (2015), 18–19.

3. D. Wu, J. Gan, Assistant haptic interaction technology for blind Internet user, J. Eng. Design, 17 (2010), 128–133.

4. X. Li, Reading resources for the blind and the present situation of reading promotion in China, New Century Lib., 5 (2013), 19–22.

5. S. Spooner, "What page, Miss?" enhancing text accessibility with DAISY (Digital Accessible Information System), J. Visual Impairment Blindness, 108 (2014), 201–211. https://doi.org/10.1177/0145482X1410800304

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. YOLOv7-WFD: A Novel Convolutional Neural Network Model for Helmet Detection in High-Risk Workplaces;IEEE Access;2023