A grapheme-based method for automatic alignment of speech and text data

Adriana Stan,Peter Bell,Simon King

doi:10.1109/slt.2012.6424237

A grapheme-based method for automatic alignment of speech and text data

Adriana Stan, Peter Bell + Show 1 more

Open Access

https://doi.org/10.1109/slt.2012.6424237

Copy DOI

Publication Date: Dec 1, 2012

Citations: 30

Affiliation: Technical University of Cluj-Napoca, University of Edinburgh

#Initial Acoustic Models #Imperfect Transcripts + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

This paper introduces a method for automatic alignment of speech data with unsynchronised, imperfect transcripts, for a domain where no initial acoustic models are available. Using grapheme-based acoustic models, word skip networks and orthographic speech transcripts, we are able to harvest 55% of the speech with a 93% utterance-level accuracy and 99% word accuracy for the produced transcriptions. The work is based on the assumption that there is a high degree of correspondence between the speech and text, and that a full transcription of all of the speech is not required. The method is language independent and the only prior knowledge and resources required are the speech and text transcripts, and a few minor user interventions.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.