Word Sense Disambiguation Using Prior Probability Estimation Based on the Korean WordNet-Reference-Cited by-同舟云学术

Word Sense Disambiguation Using Prior Probability Estimation Based on the Korean WordNet

Published:2021-11-26 Issue:23 Volume:10 Page:2938
ISSN:2079-9292
Container-title:Electronics
language:en
Short-container-title:Electronics

Author:

Kim Minho^ORCID,Kwon Hyuk-Chul

Abstract

Supervised disambiguation using a large amount of corpus data delivers better performance than other word sense disambiguation methods. However, it is not easy to construct large-scale, sense-tagged corpora since this requires high cost and time. On the other hand, implementing unsupervised disambiguation is relatively easy, although most of the efforts have not been satisfactory. A primary reason for the performance degradation of unsupervised disambiguation is that the semantic occurrence probability of ambiguous words is not available. Hence, a data deficiency problem occurs while determining the dependency between words. This paper proposes an unsupervised disambiguation method using a prior probability estimation based on the Korean WordNet. This performs better than supervised disambiguation. In the Korean WordNet, all the words have similar semantic characteristics to their related words. Thus, it is assumed that the dependency between words is the same as the dependency between their related words. This resolves the data deficiency problem by determining the dependency between words by calculating the χ2 statistic between related words. Moreover, in order to have the same effect as using the semantic occurrence probability as prior probability, which is used in supervised disambiguation, semantically related words of ambiguous vocabulary are obtained and utilized as prior probability data. An experiment was conducted with Korean, English, and Chinese to evaluate the performance of our proposed lexical disambiguation method. We found that our proposed method had better performance than supervised disambiguation methods even though our method is based on unsupervised disambiguation (using a knowledge-based approach).

Funder

Institute for Information and Communications Technology Promotion

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Computer Networks and Communications,Hardware and Architecture,Signal Processing,Control and Systems Engineering

Link

https://www.mdpi.com/2079-9292/10/23/2938/pdf

Reference33 articles.

1. Introduction to the special issue on word sense disambiguation: The state of the art;Ide;Comput. Linguist.,1998

2. Artificial intelligence based electronic healthcare solution;Kim,2021

3. Consistency of Medical Data Using Intelligent Neuron Faster R-CNN Algorithm for Smart Health Care Application

4. Word sense disambiguation

5. Applying Sentiment Product Reviews and Visualization for BI Systems in Vietnamese E-Commerce Website: Focusing on Vietnamese Context

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Unsupervised word sense disambiguation method based on word vector;2022 IEEE 2nd International Conference on Electronic Technology, Communication and Information (ICETCI);2022-05-27

2. Work of Fiction Interpretation: Corpus Approach;Filologičeskie nauki. Voprosy teorii i praktiki;2022-01