Affiliation:
1. Department of Electronics and Communication, University of Allahabad, Allahabad, India
Abstract
Karakas are an important constituent of Hindi language. Karaka relations express syntactico-semantic or semantico-syntactic relationship between verbs and nouns or pronouns in a sentence. They capture certain level of semantics closer to thematic relations, but different from it. A vibhakti is assigned to each karaka, in Paninian grammar. This paper investigates the role of karaka relations in Hindi Word Sense Disambiguation (WSD) by utilizing vibhaktis. Two supervised WSD algorithms were used for disambiguation. The first algorithm is based on conditional probability of co-occurring words and the second algorithm is Naïve Bayes (NB) classifier. The first algorithm utilizes various heuristics for analyzing the role of karakas in Hindi WSD. The authors obtained an improvement of 14.86% in precision by utilizing content words, vibhaktis and phrases containing them in context vector over the context vector of content words after dropping vibhaktis. A gain of 6.91% in precision was observed by using content words and vibhaktis in context vector over the context vector of content words after dropping vibhaktis of similar context window size. The authors obtained maximum precision of 50.73% by extracting vibhaktis in a ±3 window using WSD algorithm based on conditional probability of co-occurring words. They obtained maximum precision of 56.56% by extracting vibhaktis in a ±4 window using NB classifier.
Cited by
7 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献
1. A Novel Unsupervısed Graph-Based Algorıthm for Hindi Word Sense Disambiguation;SN Computer Science;2023-09-02
2. Comparative Analysis of Path-based Similarity Measures for Word Sense Disambiguation;2023 3rd International conference on Artificial Intelligence and Signal Processing (AISP);2023-03-18
3. Development of Automatic Rule-based Semantic Tagger and Karaka Analyzer for Hindi;ACM Transactions on Asian and Low-Resource Language Information Processing;2022-03-31
4. Nepali Word-Sense Disambiguation Using Variants of Simplified Lesk Measure;Transactions on Computer Systems and Networks;2021
5. Ambiguity Resolution : An Analytical Study;International Journal of Scientific Research in Computer Science, Engineering and Information Technology;2020-04-15