Semi-Automatic Corpus Expansion and Extraction of Uyghur-Named Entities and Relations Based on a Hybrid Method-Reference-Cited by-同舟云学术

Semi-Automatic Corpus Expansion and Extraction of Uyghur-Named Entities and Relations Based on a Hybrid Method

Published:2020-01-06 Issue:1 Volume:11 Page:31
ISSN:2078-2489
Container-title:Information
language:en
Short-container-title:Information

Author:

Halike Ayiguli^ORCID,Abiderexiti Kahaerjiang,Yibulayin Tuergen

Abstract

Relation extraction is an important task with many applications in natural language processing, such as structured knowledge extraction, knowledge graph construction, and automatic question answering system construction. However, relatively little past work has focused on the construction of the corpus and extraction of Uyghur-named entity relations, resulting in a very limited availability of relation extraction research and a deficiency of annotated relation data. This issue is addressed in the present article by proposing a hybrid Uyghur-named entity relation extraction method that combines a conditional random field model for making suggestions regarding annotation based on extracted relations with a set of rules applied by human annotators to rapidly increase the size of the Uyghur corpus. We integrate our relation extraction method into an existing annotation tool, and, with the help of human correction, we implement Uyghur relation extraction and expand the existing corpus. The effectiveness of our proposed approach is demonstrated based on experimental results by using an existing Uyghur corpus, and our method achieves a maximum weighted average between precision and recall of 61.34%. The method we proposed achieves state-of-the-art results on entity and relation extraction tasks in Uyghur.

Funder

National Natural Science Foundation of China

Publisher

MDPI AG

Subject

Information Systems

Link

https://www.mdpi.com/2078-2489/11/1/31/pdf

Reference39 articles.

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Zero-Shot Relation Triple Extraction with Prompts for Low-Resource Languages;Applied Sciences;2023-04-06

2. Bootstrapping semi-supervised annotation method for potential suicidal messages;Internet Interventions;2022-04

3. Iterative Learning for Semi-automatic Annotation Using User Feedback;Communications in Computer and Information Science;2022