Adaptive Training for Voice Conversion Based on Eigenvoices
スポンサーリンク
概要
- 論文の詳細を見る
In this paper, we describe a novel model training method for one-to-many eigenvoice conversion (EVC). One-to-many EVC is a technique for converting a specific source speakers voice into an arbitrary target speakers voice. An eigenvoice Gaussian mixture model (EV-GMM) is trained in advance using multiple parallel data sets consisting of utterance-pairs of the source speaker and many pre-stored target speakers. The EV-GMM can be adapted to new target speakers using only a few of their arbitrary utterances by estimating a small number of adaptive parameters. In the adaptation process, several parameters of the EV-GMM to be fixed for different target speakers strongly affect the conversion performance of the adapted model. In order to improve the conversion performance in one-to-many EVC, we propose an adaptive training method of the EV-GMM. In the proposed training method, both the fixed parameters and the adaptive parameters are optimized by maximizing a total likelihood function of the EV-GMMs adapted to individual pre-stored target speakers. We conducted objective and subjective evaluations to demonstrate the effectiveness of the proposed training method. The experimental results show that the proposed adaptive training yields significant quality improvements in the converted speech.
論文 | ランダム
- 内視鏡支援下に完全口内法で整復固定した関節突起基底部骨折の1例
- 経験 顎矯正手術における超音波骨メスの使用経験
- 前歯部歯槽骨切り術の臨床統計的検討
- 抗リン脂質抗体症候群を合併した顎変形症患者の全身麻酔経験
- 歯槽膿瘍に対する炭酸ガスレーザー照射後に生じた顔面・頸部・縦隔気腫の1例