X-Vectors Meet Emotions: A Study On Dependencies Between Emotion And Speaker Recognition

Raghavendra Pappagari, Tianzi Wang, JesÃºs Villalba, Nanxin Chen, Najim Dehak

DOI

SPS

Members: Free
IEEE Members: $11.00
Non-members: $15.00

Length: 16:20

04 May 2020

In this work, we explore the dependencies between speaker recognition and emotion recognition. We first show that knowledge learned for speaker recognition can be reused for emotion recognition through transfer learning. Then, we show the effect of emotion on speaker recognition. For emotion recognition, we show that using a simple linear model is enough to obtain good performance on the features extracted from pre-trained models such as the x-vector model. Then, we improve emotion recognition performance by fine-tuning for emotion classification. We evaluated our experiments on three different types of datasets: IEMOCAP, MSP-Podcast, and Crema-D. By fine-tuning, we obtained 30.40%, 7.99%, and 8.61% absolute improvement on IEMOCAP, MSP-Podcast, and Crema-D respectively over baseline model with no pre-training. Finally, we present results on the effect of emotion on speaker verification. We observed that speaker verification performance is prone to changes in test speaker emotions. We found that trials with angry utterances performed worst in all three datasets. We hope our analysis will initiate a new line of research in the speaker recognition community.

Tags:

sps conference

icassp 2020 virtual conference

May 2020

icassp 2020

X-Vectors Meet Emotions: A Study On Dependencies Between Emotion And Speaker Recognition

Raghavendra Pappagari, Tianzi Wang, JesÃºs Villalba, Nanxin Chen, Najim Dehak

Value-Added Bundle(s) Including this Product

ICASSP 2020 Virtual Conference - Presentation Videos Product Bundle

More Like This

IEEE ICASSP 2023, 4-10 June 2023, Greece. Virtual and In-Person Conference - Presentation Videos Product Bundle

IEEE ICASSP 2024, 1 4-19 April 2024, Seoul, Korea. Conference Presentation Videos Bundle

ICIP 2022, October 16-19, 2022, Bordeaux, France - Presentation Videos Product Bundle

Join the IEEE Signal Processing Society