Conferences >2018 IEEE Spoken Language Tec...

Adaptive Wavenet Vocoder for Residual Compensation in GAN-Based Voice Conversion

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

In this paper, we propose to use generative adversarial networks (GAN) together with a WaveNet vocoder to address the over-smoothing problem arising from the deep learnin...Show More

Metadata

Abstract:

In this paper, we propose to use generative adversarial networks (GAN) together with a WaveNet vocoder to address the over-smoothing problem arising from the deep learning approaches to voice conversion, and to improve the vocoding quality over the traditional vocoders. As GAN aims to minimize the divergence between the natural and converted speech parameters, it effectively alleviates the over-smoothing problem in the converted speech. On the other hand, WaveNet vocoder allows us to leverage from the human speech of a large speaker population, thus improving the naturalness of the synthetic voice. Furthermore, for the first time, we study how to use WaveNet vocoder for residual compensation to improve the voice conversion performance. The experiments show that the proposed voice conversion framework consistently outperforms the baselines.

Published in: 2018 IEEE Spoken Language Technology Workshop (SLT)

Date of Conference: 18-21 December 2018

Date Added to IEEE Xplore: 14 February 2019

ISBN Information:

DOI: 10.1109/SLT.2018.8639507

Conference Location: Athens, Greece