Speaker Segmentation Based on Subsegmental Features and Neural Network Models

N Dhananjaya,S Guruprasad,B Yegnanarayana

doi:10.1007/978-3-540-30499-9_188

Abstract

In this paper, we propose an alternate approach for detecting speaker changes in a multispeaker speech signal. Current approaches for speaker segmentation employ features based on characteristics of the vocal tract system and they rely on the dissimilarity between the distributions of two sets of feature vectors. This statistical approach to a point phenomenon (speaker change) fails when the given conversation involves short speaker turns (< 5 s duration). The excitation source signal plays an important role in characterizing a speaker’s voice. We use autoassociative neural network (AANN) models to capture the characteristics of the excitation source that are present in the linear prediction (LP) residual of speech signal. The AANN models are then used to detect the speaker changes. Results show that excitation source features provide better evidence for speaker segmentation as compared to vocal tract features.KeywordsSpeech SignalLinear PredictionVocal TractSpeaker RecognitionMiss Detection RateThese keywords were added by machine and not by the authors. This process is experimental and the keywords may be updated as the learning algorithm improves.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Speaker Segmentation Based on Subsegmental Features and Neural Network Models

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Speaker-specific information from residual phase
K Sri Rama Murty ... B Yegnanarayana
-
K Sri Rama Murty, et. al.K Sri Rama Murty ... B Yegnanarayana
11 Dec 2004
11 Dec 2004

Extraction of speaker-specific excitation information from linear prediction residual of speech
S.R Mahadeva Prasanna ... B Yegnanarayana
Speech Communication | VOL. 48
S.R Mahadeva Prasanna, et. al.S.R Mahadeva Prasanna ... B Yegnanarayana
17 Jul 2006
Speech Communication | VOL. 48

Speaker change detection in casual conversations using excitation source features
N Dhananjaya ... B Yegnanarayana
Speech Communication | VOL. 50
N Dhananjaya, et. al.N Dhananjaya ... B Yegnanarayana
07 Sep 2007
Speech Communication | VOL. 50

Speaker verification: minimizing the channel effects using autoassociative neural network models
S.P Kishore ... B Yegnanarayana
-
S.P Kishore, et. al.S.P Kishore ... B Yegnanarayana
05 Jun 2000
05 Jun 2000

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Speaker Segmentation Based on Subsegmental Features and Neural Network Models

Abstract

Talk to us

Similar Papers