Conferences >ICASSP 2019 - 2019 IEEE Inter...

Speaker Characterization Using TDNN-LSTM Based Speaker Embedding

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

In this paper we propose speaker characterization using time delay neural networks and long short-term memory neural networks (TDNN-LSTM) speaker embedding. Three types o...Show More

Metadata

Abstract:

In this paper we propose speaker characterization using time delay neural networks and long short-term memory neural networks (TDNN-LSTM) speaker embedding. Three types of front-end feature extraction are investigated to find good features for speaker embedding. Three kinds of data augmentation are used to increase the amount and diversity of the training data. The proposed methods are evaluated with the National Institute of Standards and Technology (NIST) speaker recognition evaluation (SRE) tasks. Experimental results show that the proposed methods achieve a decision cost of 0.400 with the pooled SRE 2018 development set with a single system. In addition, by applying simple average score combination on the outputs of 12 systems, the proposed methods achieve an equal error rate (EER) of 5.56% and a minimum decision cost function of 0.423 with the SRE 2016 evaluation set.

Published in: ICASSP 2019 - 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Date of Conference: 12-17 May 2019

Date Added to IEEE Xplore: 16 April 2019

ISBN Information:

ISSN Information:

DOI: 10.1109/ICASSP.2019.8683185

Conference Location: Brighton, UK

Contents

References is not available for this document.

Speaker Characterization Using TDNN-LSTM Based Speaker Embedding

Abstract:

Metadata

Abstract:

ISSN Information:

References

IEEE Account

Purchase Details

Profile Information

Need Help?

Speaker Characterization Using TDNN-LSTM Based Speaker Embedding

Alerts

Abstract:

Metadata

Abstract:

ISSN Information:

References

IEEE Account

Purchase Details

Profile Information

Need Help?