Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

Ching-Hua Chuan

doi:10.4018/jmdem.2013010101

Abstract

This paper presents an audio classification and retrieval system using wavelets for extracting low-level acoustic features. The author performed multiple-level decomposition using discrete wavelet transform to extract acoustic features from audio recordings at different scales and times. The extracted features are then translated into a compact vector representation. Gaussian mixture models with expectation maximization algorithm are used to build models for audio classes and individual audio examples. The system is evaluated using three audio classification tasks: speech/music, male/female speech, and music genre. They also show how wavelets and Gaussian mixture models are used for class-based audio retrieval in two approaches: indexing using only wavelets versus indexing by Gaussian components. By evaluating the system through 10-fold cross-validation, the author shows the promising capability of wavelets and Gaussian mixture models for audio classification and retrieval. They also compare how parameters including frame size, wavelet level, Gaussian components, and sampling size affect performance in Gaussian models.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

Abstract

Talk to us

Similar Papers

More From: International Journal of Multimedia Data Engineering and Management

Lead the way for us

Journal: International Journal of Multimedia Data Engineering and Management	Publication Date: Jan 1, 2013
Citations: 3

Similar Papers

Using Wavelets and Gaussian Mixture Models for Audio Classification
Ching-Hua Chuan ... Susan Vasana
-
Ching-Hua Chuan, et. al.Ching-Hua Chuan ... Susan Vasana
01 Dec 2012
01 Dec 2012

Content-based classification and retrieval of audio
Tong Zhang ... C.-C Jay Kuo
-
Tong Zhang, et. al.Tong Zhang ... C.-C Jay Kuo
02 Oct 1998
02 Oct 1998

Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

International Journal of Multimedia Data Engineering and Management | VOL. -

01 Jan 2013
International Journal of Multimedia Data Engineering and Management | VOL. -

Contact-state modelling in force-controlled robotic peg-in-hole assembly processes of flexible objects using optimised Gaussian mixtures
Ibrahim F Jasim ... Holger Voos
Proceedings of the Institution of Mechanical Engineers, Part B: Journal of Engineering Manufacture | VOL. 231
Ibrahim F Jasim, et. al.Ibrahim F Jasim ... Holger Voos
04 Sep 2015
Proceedings of the Institution of Mechanical Engineers, Part B: Journal of Engineering Manufacture | VOL. 231

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Audio Classification and Retrieval Using Wavelets and Gaussian Mixture Models

Abstract

Talk to us

Similar Papers

More From: International Journal of Multimedia Data Engineering and Management