Improving the Performance of Vietnamese–Korean Neural Machine Translation with Contextual Embedding-Reference-Cited by-同舟云学术

Improving the Performance of Vietnamese–Korean Neural Machine Translation with Contextual Embedding

Published:2021-11-23 Issue:23 Volume:11 Page:11119
ISSN:2076-3417
Container-title:Applied Sciences
language:en
Short-container-title:Applied Sciences

Author:

Vu Van-Hai,Nguyen Quang-Phuoc,Tunyan Ebipatei Victoria,Ock Cheol-Young

Abstract

With the recent evolution of deep learning, machine translation (MT) models and systems are being steadily improved. However, research on MT in low-resource languages such as Vietnamese and Korean is still very limited. In recent years, a state-of-the-art context-based embedding model introduced by Google, bidirectional encoder representations for transformers (BERT), has begun to appear in the neural MT (NMT) models in different ways to enhance the accuracy of MT systems. The BERT model for Vietnamese has been developed and significantly improved in natural language processing (NLP) tasks, such as part-of-speech (POS), named-entity recognition, dependency parsing, and natural language inference. Our research experimented with applying the Vietnamese BERT model to provide POS tagging and morphological analysis (MA) for Vietnamese sentences,, and applying word-sense disambiguation (WSD) for Korean sentences in our Vietnamese–Korean bilingual corpus. In the Vietnamese–Korean NMT system, with contextual embedding, the BERT model for Vietnamese is concurrently connected to both encoder layers and decoder layers in the NMT model. Experimental results assessed through BLEU, METEOR, and TER metrics show that contextual embedding significantly improves the quality of Vietnamese–Korean NMT.

Publisher

MDPI AG

Subject

Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science

Link

https://www.mdpi.com/2076-3417/11/23/11119/pdf

Reference42 articles.

1. Incorporating BERT into Neural Machine Translation;Zhu;arXiv,2020

2. ALBERT: A Lite BERT for Self-Supervised Learning of Language Representations;Lan;arXiv,2020

3. RoBERTa: A Robustly Optimized BERT Pretraining Approach;Liu;arXiv,2019

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Low-Resource Neural Machine Translation: A Systematic Literature Review;IEEE Access;2023