Cross-Corpus Speech Emotion Recognition Based on Transfer Learning and Multi-Loss Dynamic Adjustment-Reference-Cited by-同舟云学术

Cross-Corpus Speech Emotion Recognition Based on Transfer Learning and Multi-Loss Dynamic Adjustment

Published:2022-09-20 Issue: Volume:2022 Page:1-10
ISSN:1687-5273
Container-title:Computational Intelligence and Neuroscience
language:en
Short-container-title:Computational Intelligence and Neuroscience

Author:

Tao Huawei¹^ORCID,Wang Yang¹^ORCID,Zhuang Zhihao¹^ORCID,Fu Hongliang¹^ORCID,Guo Xinying¹^ORCID,Zou Shuguang¹^ORCID

Affiliation:

1. College of Information Science and Engineering, Henan University of Technology, Zhengzhou 450001, China

Abstract

In this paper, we do research on cross-corpus speech emotion recognition (SER), in which the training and testing speech signals come from different speech corpus. The mismatched feature distribution between the training and testing sets makes many classical algorithms unable to achieve better results. To deal with this issue, a transfer learning and multi-loss dynamic adjustment (TLMLDA) algorithm is initiatively proposed in this paper. The proposed algorithm first builds a novel deep network model based on a deep auto-encoder and fully connected layers to improve the representation ability of features. Subsequently, global domain and subdomain adaptive algorithms are jointly adopted to implement features transfer. Finally, dynamic weighting factors are constructed to adjust the contribution of different loss functions to prevent optimization offset of model training, which effectively improve the generalization ability of the whole system. The results of simulation experiments on Berlin, eNTERFACE, and CASIA speech corpora show that the proposed algorithm can achieve excellent recognition results, and it is competitive with most of the state-of-the-art algorithms.

Funder

National Natural Science Foundation of China

Publisher

Hindawi Limited

Subject

General Mathematics,General Medicine,General Neuroscience,General Computer Science

Link

http://downloads.hindawi.com/journals/cin/2022/5019384.pdf

Reference40 articles.

1. Self Supervised Adversarial Domain Adaptation for Cross-Corpus and Cross-Language Speech Emotion Recognition

2. Emotion recognition by speech signals;O. W. Kwon,2003

3. Emotional speech classification using Gaussian mixture models

4. Speech emotion recognition using hidden Markov models

5. Recognizing emotion in speech;F. Dellaert

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Hinglish Sentiment Analysis: Deep Learning Models for Nuanced Sentiment Classification in Multilingual Digital Communication;2024 2nd International Conference on Device Intelligence, Computing and Communication Technologies (DICCT);2024-03-15

2. Optimal Feature Learning for Speech Emotion Recognition – A DeepNet Approach;2023 International Conference on Data Science and Network Security (ICDSNS);2023-07-28