An Evaluation of Speech-Based Recognition of Emotional and Physiological Markers of Stress-Reference-Cited by-同舟云学术

An Evaluation of Speech-Based Recognition of Emotional and Physiological Markers of Stress

Published:2021-12-06 Issue: Volume:3 Page:
ISSN:2624-9898
Container-title:Frontiers in Computer Science
language:
Short-container-title:Front. Comput. Sci.

Author:

Baird Alice,Triantafyllopoulos Andreas,Zänkert Sandra,Ottl Sandra,Christ Lukas,Stappen Lukas,Konzok Julian,Sturmbauer Sarah,Meßner Eva-Maria,Kudielka Brigitte M.,Rohleder Nicolas,Baumeister Harald,Schuller Björn W.

Abstract

Life in modern societies is fast-paced and full of stress-inducing demands. The development of stress monitoring methods is a growing area of research due to the personal and economic advantages that timely detection provides. Studies have shown that speech-based features can be utilised to robustly predict several physiological markers of stress, including emotional state, continuous heart rate, and the stress hormone, cortisol. In this contribution, we extend previous works by the authors, utilising three German language corpora including more than 100 subjects undergoing a Trier Social Stress Test protocol. We present cross-corpus and transfer learning results which explore the efficacy of the speech signal to predict three physiological markers of stress—sequentially measured saliva-based cortisol, continuous heart rate as beats per minute (BPM), and continuous respiration. For this, we extract several features from audio as well as video and apply various machine learning architectures, including a temporal context-based Long Short-Term Memory Recurrent Neural Network (LSTM-RNN). For the task of predicting cortisol levels from speech, deep learning improves on results obtained by conventional support vector regression—yielding a Spearman correlation coefficient (ρ) of 0.770 and 0.698 for cortisol measurements taken 10 and 20 min after the stress period for the two corpora applicable—showing that audio features alone are sufficient for predicting cortisol, with audiovisual fusion to an extent improving such results. We also obtain a Root Mean Square Error (RMSE) of 38 and 22 BPM for continuous heart rate prediction on the two corpora where this information is available, and a normalised RMSE (NRMSE) of 0.120 for respiration prediction (−10: 10). Both of these continuous physiological signals show to be highly effective markers of stress (based on cortisol grouping analysis), both when available as ground truth and when predicted using speech. This contribution opens up new avenues for future exploration of these signals as proxies for stress in naturalistic settings.

Publisher

Frontiers Media SA

Reference83 articles.

1. Power Spectrum Analysis of Heart Rate Fluctuation: a Quantitative Probe of Beat-To-Beat Cardiovascular Control;Akselrod;Science,1981

2. Snore Sound Classification Using Image-Based Deep Spectrum Features;Amiriparian,2017

3. Using Speech to Predict Sequentially Measured Cortisol Levels during a Trier Social Stress Test;Baird,2019

4. An Evaluation of the Effect of Anxiety on Speech–Computational Prediction of Anxiety from Sustained Vowels;Baird,2020

5. A Theory of Learning from Different Domains;Ben-David;Mach Learn.,2010

Cited by 15 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Fear of falling in community-dwelling older adults: What their gait acceleration pattern reveals;Computer Methods and Programs in Biomedicine;2024-02

2. Ecologically valid speech collection in behavioral research: The Ghent Semi-spontaneous Speech Paradigm (GSSP);Behavior Research Methods;2023-12-13

3. Investigating the Generalizability of Physiological Characteristics of Anxiety;2023 IEEE International Conference on Bioinformatics and Biomedicine (BIBM);2023-12-05

4. VoStress – Voice-based Detection of Acute Psychosocial Stress;2023 IEEE EMBS International Conference on Biomedical and Health Informatics (BHI);2023-10-15

5. HEAR4Health: a blueprint for making computer audition a staple of modern healthcare;Frontiers in Digital Health;2023-09-12