Abstract
In this paper we have analysed emotion recognition performance in speaker dependent, text dependent, text independent, speaker independent, language dependent and cross language emotion recognition from speech. These studies were carried out using Gaussian Mixture Model (GMM) and Hidden Markov Model (HMM) as classification models. IITKGP-SESC and IITKGP-SEHSC emotional speech corpora are used for carried out these studies. The emotions considered in this study are anger, disgust, fear, happy, neutral, sarcastic, and surprise. Mel Frequency Cepstral Coefficients (MFCCs) features are used for identifying the emotions. Emotion recognition performance of speaker dependent mode is better than speaker independent and cross language modes. From the results it is observed that emotion recognition performance depends on the speaker and language.
Talk to us
Join us for a 30 min session where you can share your feedback and ask us any queries you have
Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.