Using Voice-quality Measurements with Prosodic and Spectral Features for Speaker Diarization

Abraham Woubie ,Javier Hernando ,Jordi Luque

doi:10.5281/zenodo.31035

Using Voice-quality Measurements with Prosodic and Spectral Features for Speaker Diarization

Abraham Woubie , Javier Hernando + Show 1 more

https://doi.org/10.5281/zenodo.31035

Copy DOI

Publication Date: Apr 5, 2016
Citations: 19	License type: cc-by-nc-nd

Affiliation: Universitat Politècnica de Catalunya, Telefonica Research and Development, Clínica Diagonal

#Augmented Multi-party Interaction #Short-term Spectral Features + Show 8 more

Abstract
Full-Text
Similar Papers

Abstract

Jitter and shimmer voice-quality measurements have been successfully used to detect voice pathologies and classify different speaking styles. In this paper, we investigate the usefulness of jitter and shimmer voice measurements in the framework of the speaker diarization task. The combination of jitter and shimmer voice-quality features with the long-term prosodic and short-term spectral features is explored in a subset of the Augmented Multi-party Interaction (AMI) corpus, a multi-party and spontaneous speech set of recordings. The best results have been obtained by fusing the voice-quality features with the prosodic ones at the feature level, and then fusing them with the spectral features at the score level. Experimental results show more than 20% relative DER improvement compared to the spectral baseline system.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.