A Deep Learning Approach for Voice Disorder Detection for Smart Connected Living Environments-Reference-Cited by-同舟云学术

A Deep Learning Approach for Voice Disorder Detection for Smart Connected Living Environments

Published:2022-02-28 Issue:1 Volume:22 Page:1-16
ISSN:1533-5399
Container-title:ACM Transactions on Internet Technology
language:en
Short-container-title:ACM Trans. Internet Technol.

Author:

Verde Laura¹,Brancati Nadia¹,De Pietro Giuseppe¹,Frucci Maria¹,Sannino Giovanna¹^ORCID

Affiliation:

1. Institute of High-Performance Computing and Networking (ICAR)—National Research Council of Italy (CNR), Naples, Italy

Abstract

Edge Analytics and Artificial Intelligence are important features of the current smart connected living community. In a society where people, homes, cities, and workplaces are simultaneously connected through various devices, primarily through mobile devices, a considerable amount of data is exchanged, and the processing and storage of these data are laborious and difficult tasks. Edge Analytics allows the collection and analysis of such data on mobile devices, such as smartphones and tablets, without involving any cloud-centred architecture that cannot guarantee real-time responsiveness. Meanwhile, Artificial Intelligence techniques can constitute a valid instrument to process data, limiting the computation time, and optimising decisional processes and predictions in several sectors, such as healthcare. Within this field, in this article, an approach able to evaluate the voice quality condition is proposed. A fully automatic algorithm, based on Deep Learning, classifies a voice as healthy or pathological by analysing spectrogram images extracted by means of the recording of vowel /a/, in compliance with the traditional medical protocol. A light Convolutional Neural Network is embedded in a mobile health application in order to provide an instrument capable of assessing voice disorders in a fast, easy, and portable way. Thus, a straightforward mobile device becomes a screening tool useful for the early diagnosis, monitoring, and treatment of voice disorders. The proposed approach has been tested on a broad set of voice samples, not limited to the most common voice diseases but including all the pathologies present in three different databases achieving F1-scores, over the testing set, equal to 80%, 90%, and 73%. Although the proposed network consists of a reduced number of layers, the results are very competitive compared to those of other “cutting edge” approaches constructed using more complex neural networks, and compared to the classic deep neural networks, for example, VGG-16 and ResNet-50.

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Networks and Communications

Link

https://dl.acm.org/doi/pdf/10.1145/3433993

Reference56 articles.

1. Voice pathology detection and classification using auto-correlation and entropy features in different frequency regions;Al-Nasheri Ahmed;IEEE Access,2017

2. Voice pathology detection using deep learning on mobile healthcare framework;Alhussein Musaed;IEEE Access,2018

3. Intelligent pathological voice detection;Ali Akbar;International Journal of Innovative Research in Technology,2018

4. Detecting Parkinson's disease with sustained phonation and speech signals using machine learning techniques;Almeida Jefferson S.;Pattern Recognition Letters,2019

Cited by 12 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Diagnosis of pathological speech with streamlined features for long short-term memory learning;Computers in Biology and Medicine;2024-03

2. A depthwise separable CNN-based interpretable feature extraction network for automatic pathological voice detection;Biomedical Signal Processing and Control;2024-02

3. Automatic Speech and Voice Disorder Detection Using Deep Learning—A Systematic Literature Review;IEEE Access;2024

4. Toward a lightweight ASR solution for atypical speech on the edge;Future Generation Computer Systems;2023-12

5. Applications of edge analytics: a systematic review;Acta Universitatis Sapientiae, Informatica;2023-12-01