Neural Sign Language Translation Based on Human Keypoint Estimation-Reference-Cited by-同舟云学术

Neural Sign Language Translation Based on Human Keypoint Estimation

Published:2019-07-01 Issue:13 Volume:9 Page:2683
ISSN:2076-3417
Container-title:Applied Sciences
language:en
Short-container-title:Applied Sciences

Author:

Ko Sang-Ki^ORCID,Kim Chang Jo^ORCID,Jung Hyedong,Cho Choongsang

Abstract

We propose a sign language translation system based on human keypoint estimation. It is well-known that many problems in the field of computer vision require a massive dataset to train deep neural network models. The situation is even worse when it comes to the sign language translation problem as it is far more difficult to collect high-quality training data. In this paper, we introduce the KETI (Korea Electronics Technology Institute) sign language dataset, which consists of 14,672 videos of high resolution and quality. Considering the fact that each country has a different and unique sign language, the KETI sign language dataset can be the starting point for further research on the Korean sign language translation. Using the KETI sign language dataset, we develop a neural network model for translating sign videos into natural language sentences by utilizing the human keypoints extracted from the face, hands, and body parts. The obtained human keypoint vector is normalized by the mean and standard deviation of the keypoints and used as input to our translation model based on the sequence-to-sequence architecture. As a result, we show that our approach is robust even when the size of the training data is not sufficient. Our translation model achieved 93.28% (55.28%, respectively) translation accuracy on the validation set (test set, respectively) for 105 sentences that can be used in emergency situations. We compared several types of our neural sign translation models based on different attention mechanisms in terms of classical metrics for measuring the translation performance.

Funder

Institute for Information and communications Technology Promotion

Publisher

MDPI AG

Subject

Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science

Link

https://www.mdpi.com/2076-3417/9/13/2683/pdf

Reference59 articles.

1. Long Short-Term Memory

Cited by 103 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Reviewing 25 years of continuous sign language recognition research: Advances, challenges, and prospects;Information Processing & Management;2024-09

2. From rule-based models to deep learning transformers architectures for natural language processing and sign language translation systems: survey, taxonomy and performance evaluation;Artificial Intelligence Review;2024-08-29

3. Techniques for Generating Sign Language a Comprehensive Review;Journal of The Institution of Engineers (India): Series B;2024-07-13

4. A survey on recent advances in Sign Language Production;Expert Systems with Applications;2024-06

5. SynthSL: Expressive Humans for Sign Language Image Synthesis;2024 IEEE 18th International Conference on Automatic Face and Gesture Recognition (FG);2024-05-27