Recognition of Transient Environmental Sounds Based on Temporal and Frequency Features-Reference-Cited by-同舟云学术

Recognition of Transient Environmental Sounds Based on Temporal and Frequency Features

Published:2019-11-05 Issue:6 Volume:13 Page:803-809
ISSN:1883-8022
Container-title:International Journal of Automation Technology
language:en
Short-container-title:IJAT

Author:

Okubo Shota, ,Gong Zhihao,Fujita Kento,Sasaki Ken

Abstract

Environmental sound recognition (ESR) refers to the recognition of all sounds other than the human voice or musical sounds. Typical ESR methods utilize spectral information and variation within it with respect to time. However, in the case of transient sounds, spectral information is insufficient because only an average quantity of a given signal within a time period can be recognized. In this study, the waveform of sound signals and their spectrum were analyzed visually to extract temporal characteristics of the sound more directly. Based on the observations, features such as the initial rise time, duration, and smoothness of the sound signal; the distribution and smoothness of the spectrum; the clarity of the sustaining sound components; and the number and interval of collisions in chattering were proposed. Experimental feature values were obtained for eight transient environmental sounds, and the distributions of the values were evaluated. A recognition experiment was conducted on 11 transient sounds. The Mel-frequency cepstral coefficient (MFCC) was selected as reference. A support vector machine was adopted as the classification algorithm. The recognition rates obtained from the MFCC were below 50% for five of the 11 sounds, and the overall recognition rate was 69%. In contrast, the recognition rates obtained using the proposed features were above 50% for all sounds, and the overall rate was 86%.

Publisher

Fuji Technology Press Ltd.

Subject

Industrial and Manufacturing Engineering,Mechanical Engineering

Reference27 articles.

1. A. Caggiano and L. Nele, “Artificial Neural Networks for Tool Wear Prediction Based on Sensor Fusion Monitoring of CFRP/CFRP Stack Drilling,” Int. J. Automation Technol., Vol.12, No.3, pp. 275-281, 2018.

2. T. Zhang, Z. M. Zeng, Y. B. Li, W. K. Wang, and X. Bian, “Characteristics Analysis of Vacuum Gas Leak Detection Signals Based on Acoustic Emission,” Int. J. Automation Technol., Vol.8, No.1, pp. 57-61, 2014.

3. G. Gu, R. Hu, and Y. Li, “Study on Identification of Damage to Wind Turbine Blade Based on Support Vector Machine and Particle Swarm Optimization,” J. Robot. Mechatron., Vol.27, No.3, pp. 244-250, 2015.

4. H. Kalantarian, N. Alshurafa, M. Pourhomayoun, S. Sarin, T. Le, and M. Sarrafzadeh, “Spectrogram-Based Audio Classification of Nutrition Intake,” 2014 IEEE Health Innovations and Point-of-Care Technologies Conf. (HIC 2014), pp. 161-164, 2014.

5. A. Kamiyanagi, Y. Sumita, M. Chikai, K. Kimura, Y. Seki, S. Ino, and H. Taniguchi, “Evaluation of Swallowing Sound Using a Throat Microphone with an AE Sensor in Patients Wearing Palatal Augmentation Prosthesis,” J. Adv. Comput. Intell. Intell. Inform., Vol.21, No.3, pp. 573-580, 2017.

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Effect of FDM Processing Conditions on Snap-Fit Characteristic in Assembly;International Journal of Automation Technology;2023-07-05

2. Tunnel abnormal sound recognition based on multi-channel convolutional neural network;2022 18th International Conference on Computational Intelligence and Security (CIS);2022-12