Implementation of a System for Assessing the Quality of Spoken English Pronunciation Based on Cognitive Heuristic Computing-Reference-Cited by-同舟云学术

Implementation of a System for Assessing the Quality of Spoken English Pronunciation Based on Cognitive Heuristic Computing

Published:2022-07-08 Issue: Volume:2022 Page:1-12
ISSN:1687-5273
Container-title:Computational Intelligence and Neuroscience
language:en
Short-container-title:Computational Intelligence and Neuroscience

Author:

Wu Yanping¹^ORCID,Zheng Changlong²^ORCID,Hao Meihui³^ORCID,Wang Linlin³^ORCID

Affiliation:

1. Department of Economic Management, Dongchang College of Liaocheng University, Liaocheng, Shandong 252000, China

2. Liaocheng Yucai School, Liaocheng, Shandong 252000, China

3. Dongchang Middle School of Liaocheng Economic and Technological Development Zone, Liaocheng, Shandong 252000, China

Abstract

This paper analyzes and investigates the quality assessment of spoken English pronunciation using a cognitive heuristic computing approach and designs a corresponding spoken pronunciation quality assessment system for practical training. Using the general Goodness of Pronunciation assessment algorithm as a benchmark, the shortcomings of the traditional Goodness of Pronunciation method are explored through statistical experiments, and the validity of the overall posterior probability output from the speech model for pronunciation quality assessment is verified. For the analysis of rhythm, there is no common algorithm framework, but in this paper, the F0 similarity algorithm based on dynamic time regularization and the stop similarity algorithm based on forced alignment is proposed for the two main factors of rhythm, intonation, and pause, respectively. After framing, the Hamming window processing is used to make the signal smoother, reduce the side lobe size after fast Fourier transform processing, and solve the problem of spectrum leakage. Compared with the ordinary rectangular window function, the Hamming window can obtain a higher quality spectrum. And combined with CTC for speech recognition modeling, the recognition rates are comparable in the case of using BLSTM and bidirectional threshold cyclic unit BGRU as the hidden layer unit, respectively, and the training time is 23% less than BLSTM using BGRU; in addition, the BGRU-CTC model is improved by using a 2-BGRU-CTC model with 256 hidden layer nodes, so that the error rate of phoneme recognition is reduced to 33%. The effectiveness of the algorithm framework is also verified through experiments, which further proves the effectiveness of our proposed phoneme segment feature and rhyme similarity algorithm.

Publisher

Hindawi Limited

Subject

General Mathematics,General Medicine,General Neuroscience,General Computer Science

Link

http://downloads.hindawi.com/journals/cin/2022/5239375.pdf

Reference20 articles.

1. Quality evaluation of English pronunciation based on artificial emotion recognition and gaussian mixture model

2. Editorial for EAIT issue 2, 2019

3. The Design Patterns for Language Learning and the Assessment on Game-Based Learning

4. Computer aided reading and pronunciation practice system for elementary level: development and usability[J];R. Catanghal;Asian Journal of Multidisciplinary Studies,2018

5. A mobile game-based learning system for diacritic insertion

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. A Method for Improving the Pronunciation Quality of Vocal Music Students Based on Big Data Technology;International Journal of Web-Based Learning and Teaching Technologies;2023-12-18

2. Retracted: Implementation of a System for Assessing the Quality of Spoken English Pronunciation Based on Cognitive Heuristic Computing;Computational Intelligence and Neuroscience;2023-07-26