Text Separation From Document Images-Reference-Cited by-同舟云学术

Text Separation From Document Images

Published:2020 Issue: Volume: Page:283-313
ISSN:2327-039X
Container-title:Machine Learning and Deep Learning in Real-Time Applications
language:
Short-container-title:

Author:

Rege Priti P.¹^ORCID,Akhter Shaheera²^ORCID

Affiliation:

1. College of Engineering, Pune, India

2. Government College of Engineering, Pune, India

Abstract

Text separation in document image analysis is an important preprocessing step before executing an optical character recognition (OCR) task. It is necessary to improve the accuracy of an OCR system. Traditionally, for separating text from a document, different feature extraction processes have been used that require handcrafting of the features. However, deep learning-based methods are excellent feature extractors that learn features from the training data automatically. Deep learning gives state-of-the-art results on various computer vision, image classification, segmentation, image captioning, object detection, and recognition tasks. This chapter compares various traditional as well as deep-learning techniques and uses a semantic segmentation method for separating text from Devanagari document images using U-Net and ResU-Net models. These models are further fine-tuned for transfer learning to get more precise results. The final results show that deep learning methods give more accurate results compared with conventional methods of image processing for Devanagari text extraction.

Publisher

IGI Global

Reference42 articles.

1. Afshine Amidi and Shervine Amidi. (2018). Deep Learning cheatsheet [CS 229 - Machine Learning]. Retrieved from https://stanford.edu/~shervine/teaching/cs-229/cheatsheet-deep-learning

2. PAGE SEGMENTATION AND CLASSIFICATION UTILIZING BOTTOM-UP APPROACH

3. Antonacopoulos, A., & Ritchings, R. T. (1995). Segmentation and classification of document images. Academic Press.

4. SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation

5. Line and Word Segmentation of Arabic handwritten documents using;A.Belabiod;Neural Networks,2018

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Embedded System Based Raspberry Pi 4 for Text Detection and Recognition;Intelligent Automation & Soft Computing;2023

2. Multi-task learning for pre-processing of printed Devanagari document images with hyper-parameter optimization of the deep architecture using Taguchi method;Sādhanā;2021-07-26