Adaptive Training for Voice Conversion Based on Eigenvoices

Yamato Ohtani,Kiyohiro Shikano,Hiroshi Saruwatari,Tomoki Toda

doi:10.1587/transinf.e93.d.1589

Yamato Ohtani, Kiyohiro Shikano + Show 2 more

Open Access

https://doi.org/10.1587/transinf.e93.d.1589

Copy DOI

Abstract

In this paper, we describe a novel model training method for one-to-many eigenvoice conversion (EVC). One-to-many EVC is a technique for converting a specific source speaker's voice into an arbitrary target speaker's voice. An eigenvoice Gaussian mixture model (EV-GMM) is trained in advance using multiple parallel data sets consisting of utterance-pairs of the source speaker and many pre-stored target speakers. The EV-GMM can be adapted to new target speakers using only a few of their arbitrary utterances by estimating a small number of adaptive parameters. In the adaptation process, several parameters of the EV-GMM to be fixed for different target speakers strongly affect the conversion performance of the adapted model. In order to improve the conversion performance in one-to-many EVC, we propose an adaptive training method of the EV-GMM. In the proposed training method, both the fixed parameters and the adaptive parameters are optimized by maximizing a total likelihood function of the EV-GMMs adapted to individual pre-stored target speakers. We conducted objective and subjective evaluations to demonstrate the effectiveness of the proposed training method. The experimental results show that the proposed adaptive training yields significant quality improvements in the converted speech.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: IEICE Transactions on Information and Systems	Publication Date: Jan 1, 2010
Citations: 24	License type: free

R Discovery Prime

R Discovery Prime

Adaptive Training for Voice Conversion Based on Eigenvoices

Abstract

Talk to us

Similar Papers

More From: IEICE Transactions on Information and Systems

Lead the way for us

Similar Papers

Effective Emotion Transplantation in an End-to-End Text-to-Speech System
Young-Sun Joo ... Young-Ik Kim
IEEE Access | VOL. 8
Young-Sun Joo, et. al.Young-Sun Joo ... Young-Ik Kim
01 Jan 2020
IEEE Access | VOL. 8

One-Shot Voice Conversion Algorithm Based on Representations Separation
Chunhui Deng ... Ying Chen
IEEE Access | VOL. 8
Chunhui Deng, et. al.Chunhui Deng ... Ying Chen
01 Jan 2020
IEEE Access | VOL. 8

Voice Transformation by Mapping the Features at Syllable Level
K Sreenivasa Rao ... R H Laskar
-
K Sreenivasa Rao, et. al.K Sreenivasa Rao ... R H Laskar
18 Dec 2007
18 Dec 2007

Improvements of the One-to-Many Eigenvoice Conversion System
Yamato Ohtani ... Kiyohiro Shikano
IEICE Transactions on Information and Systems | VOL. E93-D
Yamato Ohtani, et. al.Yamato Ohtani ... Kiyohiro Shikano
01 Jan 2009
IEICE Transactions on Information and Systems | VOL. E93-D

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Adaptive Training for Voice Conversion Based on Eigenvoices

Abstract

Talk to us

Similar Papers

More From: IEICE Transactions on Information and Systems