Speaker state classification based on fusion of asymmetric simple partial least squares (SIMPLS) and support vector machines

Dong-Yan Huang,Zhengchen Zhang,Shuzhi Sam Ge

doi:10.1016/j.csl.2013.06.002

Abstract

This paper presents our studies of the effects of acoustic features, speaker normalization methods, and statistical modeling techniques on speaker state classification. We focus on the investigation of the effect of simple partial least squares (SIMPLS) in unbalanced binary classification. Beyond dimension reduction and low computational complexity, SIMPLS classifier (SIMPLSC) shows, especially, higher prediction accuracy to the class with the smaller data number. Therefore, an asymmetric SIMPLS classifier (ASIMPLSC) is proposed to enhance the performance of SIMPLSC to the class with the larger data number. Furthermore, we combine multiple system outputs (ASIMPLS classifier and Support Vector Machines) by score-level fusion to exploit the complementary information in diverse systems. The proposed speaker state classification system is evaluated with several experiments on unbalanced data sets. Within the Interspeech 2011 Speaker State Challenge, we could achieve the best results for the 2-class task of the Sleepiness Sub-Challenge with an unweighted average recall of 71.7%. Further experimental results on the SEMAINE data sets show that the ASIMPLSC achieves an absolute improvement of 6.1%, 6.1%, 24.5%, and 1.3% on the weighted average recall value, over the AVEC 2011 baseline system on the emotional speech binary classification tasks of four dimensions, namely, activation, expectation, power, and valence, respectively.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Speaker state classification based on fusion of asymmetric simple partial least squares (SIMPLS) and support vector machines

Abstract

Talk to us

Similar Papers

More From: Computer Speech & Language

Lead the way for us

Journal: Computer Speech & Language	Publication Date: Jun 25, 2013
Citations: 26

Similar Papers

Fast regression methods in a Lanczos (or PLS-1) basis. Theory and applications
Wen Wu ... Rolf Manne
Chemometrics and Intelligent Laboratory Systems | VOL. 51
Wen Wu, et. al.Wen Wu ... Rolf Manne
01 Jul 2000
Chemometrics and Intelligent Laboratory Systems | VOL. 51

Speaker state classification based on fusion of asymmetric SIMPLS and support vector machines
Dong-Yan Huang ... Zhengchen Zhang
-
Dong-Yan Huang, et. al.Dong-Yan Huang ... Zhengchen Zhang
27 Aug 2011
27 Aug 2011

Partial functional linear quantile regression for neuroimaging data analysis
Dengdeng Yu ... Ivan Mizera
Neurocomputing | VOL. 195
Dengdeng Yu, et. al.Dengdeng Yu ... Ivan Mizera
04 Feb 2016
Neurocomputing | VOL. 195

Using Multiple SVM Models for Unbalanced Credit Scoring Data Sets
Klaus B Schebesch ... Ralf Stecking
-
Klaus B Schebesch, et. al.Klaus B Schebesch ... Ralf Stecking
01 Jan 2008
01 Jan 2008

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Speaker state classification based on fusion of asymmetric simple partial least squares (SIMPLS) and support vector machines

Abstract

Talk to us

Similar Papers

More From: Computer Speech &amp; Language

More From: Computer Speech & Language