Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models-Reference-Cited by-同舟云学术

Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

Published:2013-01 Issue:1 Volume:4 Page:1-20
ISSN:1947-8534
Container-title:International Journal of Multimedia Data Engineering and Management
language:en
Short-container-title:

Author:

Chuan Ching-Hua¹

Affiliation:

1. School of Computing, University of North Florida, Jacksonville, FL, USA

Abstract

This paper presents an audio classification and retrieval system using wavelets for extracting low-level acoustic features. The author performed multiple-level decomposition using discrete wavelet transform to extract acoustic features from audio recordings at different scales and times. The extracted features are then translated into a compact vector representation. Gaussian mixture models with expectation maximization algorithm are used to build models for audio classes and individual audio examples. The system is evaluated using three audio classification tasks: speech/music, male/female speech, and music genre. They also show how wavelets and Gaussian mixture models are used for class-based audio retrieval in two approaches: indexing using only wavelets versus indexing by Gaussian components. By evaluating the system through 10-fold cross-validation, the author shows the promising capability of wavelets and Gaussian mixture models for audio classification and retrieval. They also compare how parameters including frame size, wavelet level, Gaussian components, and sampling size affect performance in Gaussian models.

Publisher

IGI Global

Reference30 articles.

1. Improving timbre similarity: How high’s the sky?;J.-J.Aucouturier;Journal of Negative Results in Speech and Audio Sciences,2004

2. Chandrasekhar, V., Sharifi, M., & Ross, D. A. (2011) Survey and evaluation of audio fingerprinting schemes for mobile query-by-example applications. In Klapuri, A., & Leider, C (Eds.), Proceedings of the 12th International Society for Music Information Retrieval Conference, Miami, FL (pp. 801-806).

3. A Noise-Robust FFT-Based Auditory Spectrum With Application in Audio Classification

4. Ten Lectures on Wavelets

5. Classification of audio signals using SVM and RBFNN

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. An Overview of Audio Event Detection Methods from Feature Extraction to Classification;Applied Artificial Intelligence;2017-11-26