Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

Ching-Hua Chuan

Source Title: International Journal of Multimedia Data Engineering and Management (IJMDEM)4(1)

ISSN: 1947-8534|EISSN: 1947-8542|EISBN13: 9781466630741|DOI: 10.4018/jmdem.2013010101

MLA

Chuan, Ching-Hua. "Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models." IJMDEM vol.4, no.1 2013: pp.1-20. http://doi.org/10.4018/jmdem.2013010101

APA

Chuan, C. (2013). Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models. International Journal of Multimedia Data Engineering and Management (IJMDEM), 4(1), 1-20. http://doi.org/10.4018/jmdem.2013010101

Chicago

Chuan, Ching-Hua. "Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models," International Journal of Multimedia Data Engineering and Management (IJMDEM) 4, no.1: 1-20. http://doi.org/10.4018/jmdem.2013010101

Export Reference

Favorite Full-Issue Download

View Full Text HTML

View Full Text PDF

Abstract

This paper presents an audio classification and retrieval system using wavelets for extracting low-level acoustic features. The author performed multiple-level decomposition using discrete wavelet transform to extract acoustic features from audio recordings at different scales and times. The extracted features are then translated into a compact vector representation. Gaussian mixture models with expectation maximization algorithm are used to build models for audio classes and individual audio examples. The system is evaluated using three audio classification tasks: speech/music, male/female speech, and music genre. They also show how wavelets and Gaussian mixture models are used for class-based audio retrieval in two approaches: indexing using only wavelets versus indexing by Gaussian components. By evaluating the system through 10-fold cross-validation, the author shows the promising capability of wavelets and Gaussian mixture models for audio classification and retrieval. They also compare how parameters including frame size, wavelet level, Gaussian components, and sampling size affect performance in Gaussian models.

You do not own this content. Please login to recommend this title to your institution's librarian or purchase it from the IGI Global bookstore.

Username or email: *

Password: *

Forgot individual login password?

Create individual account

Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

MLA

APA

Chicago

Export Reference

Abstract

Request Access