Conferences >2018 IEEE Spoken Language Tec...

An Evaluation of Deep Spectral Mappings and WaveNet Vocoder for Voice Conversion

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

This paper presents an evaluation of deep spectral mapping and WaveNet vocoder in voice conversion (VC). In our VC framework, spectral features of an input speaker are co...Show More

Metadata

Abstract:

This paper presents an evaluation of deep spectral mapping and WaveNet vocoder in voice conversion (VC). In our VC framework, spectral features of an input speaker are converted into those of a target speaker using the deep spectral mapping, and then together with the excitation features, the converted waveform is generated using WaveNet vocoder. In this work, we compare three different deep spectral mapping networks, i.e., a deep single density network (DSDN), a deep mixture density network (DMDN), and a long short-term memory recurrent neural network with an autoregressive output layer (LSTM-AR). Moreover, we also investigate several methods for reducing mismatches of spectral features used in WaveNet vocoder between training and conversion processes, such as some methods to alleviate oversmoothing effects of the converted spectral features, and another method to refine WaveNet using the converted spectral features. The experimental results demonstrate that the LSTM-AR yields nearly better spectral mapping accuracy than the others, and the proposed WaveNet refinement method significantly improves the naturalness of the converted waveform.

Published in: 2018 IEEE Spoken Language Technology Workshop (SLT)

Date of Conference: 18-21 December 2018

Date Added to IEEE Xplore: 14 February 2019

ISBN Information:

DOI: 10.1109/SLT.2018.8639608

Conference Location: Athens, Greece

Contents

References is not available for this document.

An Evaluation of Deep Spectral Mappings and WaveNet Vocoder for Voice Conversion

Abstract:

Metadata

Abstract:

References

IEEE Account

Purchase Details

Profile Information

Need Help?

An Evaluation of Deep Spectral Mappings and WaveNet Vocoder for Voice Conversion

Alerts

Abstract:

Metadata

Abstract:

References

IEEE Account

Purchase Details

Profile Information

Need Help?