Phonetic Feedback For Speech Enhancement With And Without Parallel Speech Data

Peter Plantinga, Deblin Bagchi, Eric Fosler-Lussier

DOI

SPS

Members: Free
IEEE Members: $11.00
Non-members: $15.00

Length: 14:40

04 May 2020

While deep learning systems have gained significant ground in speech enhancement research, these systems have yet to make use of the full potential of deep learning systems to provide high-level feedback. In particular, phonetic feedback is rare in speech enhancement research even though it includes valuable top-down information. We use the technique of mimic loss to provide phonetic feedback to an off-the-shelf enhancement system, and find gains in objective intelligibility scores on CHiME-4 data. This technique takes a frozen acoustic model trained on clean speech to provide valuable feedback to the enhancement model, even in the case where no parallel speech data is available. Our work is one of the first to show intelligibility improvement for neural enhancement systems without parallel speech data, and we show phonetic feedback can improve a state-of-the-art neural enhancement system trained with parallel speech data.

Tags:

sps conference

icassp 2020 virtual conference

May 2020

icassp 2020

Phonetic Feedback For Speech Enhancement With And Without Parallel Speech Data

Peter Plantinga, Deblin Bagchi, Eric Fosler-Lussier

Value-Added Bundle(s) Including this Product

ICASSP 2020 Virtual Conference - Presentation Videos Product Bundle

More Like This

IEEE ICASSP 2023, 4-10 June 2023, Greece. Virtual and In-Person Conference - Presentation Videos Product Bundle

IEEE ICASSP 2024, 1 4-19 April 2024, Seoul, Korea. Conference Presentation Videos Bundle

ICIP 2022, October 16-19, 2022, Bordeaux, France - Presentation Videos Product Bundle

Join the IEEE Signal Processing Society