Adaptive Training for Voice Conversion Based on Eigenvoices

OHTANI Yamato, TODA Tomoki, SARUWATARI Hiroshi, SHIKANO Kiyohiro

doi:10.1587/transinf.e93.d.1589

Abstract

In this paper, we describe a novel model training method for one-to-many eigenvoice conversion (EVC). One-to-many EVC is a technique for converting a specific source speaker's voice into an arbitrary target speaker's voice. An eigenvoice Gaussian mixture model (EV-GMM) is trained in advance using multiple parallel data sets consisting of utterance-pairs of the source speaker and many pre-stored target speakers. The EV-GMM can be adapted to new target speakers using only a few of their arbitrary utterances by estimating a small number of adaptive parameters. In the adaptation process, several parameters of the EV-GMM to be fixed for different target speakers strongly affect the conversion performance of the adapted model. In order to improve the conversion performance in one-to-many EVC, we propose an adaptive training method of the EV-GMM. In the proposed training method, both the fixed parameters and the adaptive parameters are optimized by maximizing a total likelihood function of the EV-GMMs adapted to individual pre-stored target speakers. We conducted objective and subjective evaluations to demonstrate the effectiveness of the proposed training method. The experimental results show that the proposed adaptive training yields significant quality improvements in the converted speech.

Journal

IEICE Transactions on Information and Systems

IEICE Transactions on Information and Systems E93-D (6), 1589-1598, 2010

The Institute of Electronics, Information and Communication Engineers

Keywords

Details 詳細情報について

CRID: 1390282679356320256

NII Article ID: 120005716790; 10027988002

NII Book ID: AA10826272

DOI: 10.1587/transinf.e93.d.1589

ISSN: 17451361; 09168532

HANDLE: 10061/7843

Web Site: https://naist.repo.nii.ac.jp/records/3851; http://www.jstage.jst.go.jp/article/transinf/E93.D/6/E93.D_6_1589/_pdf

Text Lang: en

Data Source

JaLC
IRDB
Crossref
CiNii Articles
KAKEN

Abstract License Flag: Disallowed

Export

Adaptive Training for Voice Conversion Based on Eigenvoices

Search this article

Abstract

Journal

Citations (6)*help

References(18)*help

Related Projects

Keywords

Details 詳細情報について

Export

Report a problem

Adaptive Training for Voice Conversion Based on Eigenvoices

Search this article

Abstract

Journal

Citations (6)*help

References(18)*help

Related Projects

Keywords

Details 詳細情報について

Export

Report a problem

Project list