The Effects of OCR Error on the Extraction of Private Information

Author:

Taghva Kazem,Beckley Russell,Coombs Jeffrey

Publisher

Springer Berlin Heidelberg

Reference18 articles.

1. U.S. Government. The freedom of information act 5 U.S.C. sec. 552 as amended in 2002 (Viewed June 30 2004), http://www.usdoj.gov/oip/foia_updates/Vol_XVII_4/page2.htm

2. U.S. Government. Frequently occurring first names and surnames from the 1990 census (Viewed August, 2005), http://www.census.gov/genealogy/www/freqnames.html

3. Lecture Notes in Computer Science;R. Grishman,1997

4. Jing, H., Lopresti, D., Shih, C.: Summarizing noisy documents. In: Proceedings of SDIUT 2003, Greenbelt, MD, April 2003, pp. 111–119 (2003)

5. McCallum, A.: Bow: A toolkit for statistical language modeling, text retrieval, classification and clustering (1996), http://www.cs.cmu.edu/~mccallum/bow

Cited by 11 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

1. Deep learning approaches for information extraction from visually rich documents: datasets, challenges and methods;International Journal on Document Analysis and Recognition (IJDAR);2024-07-29

2. Zone and rule assisted recognition of Meitei-Mayek handwritten characters;Evolutionary Intelligence;2024-03-21

3. OCR-Free Document Understanding Transformer;Lecture Notes in Computer Science;2022

4. Global Postal Automation;Lecture Notes in Networks and Systems;2021-08-07

5. Name identification and extraction with formal concept analysis;International Journal of Machine Learning and Cybernetics;2016-03-18

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3