Phoneme-Based Distribution Regularization For Speech Enhancement

Yajing Liu, Xiulian Peng, Zhiwei Xiong, Yan Lu

DOI

10.17023/8xxa-2057

SPS

Members: Free
IEEE Members: $11.00
Non-members: $15.00

Length: 00:10:40

10 Jun 2021

Existing speech enhancement methods mainly focus on the signal level similarity of the enhanced speech and the target. They do not pay attention to understanding the whole speech and context. Therefore, the recognizability and coherence of the enhanced speech are impaired. To address this problem, we propose a phoneme-based distribution regularization (PbDr) for speech enhancement, which aims to incorporate context information into speech enhancement network in a conditional manner to achieve better perceptual quality and better recognizability. As different phonemes always lead to different feature distributions in frequency, we propose to learn a parameter pair, i.e. scale and bias, through a phoneme classification vector to modulate the speech enhancement network. The modulation parameter pair not only includes frame-wise condition but also frequency-wise condition, which effectively map features to phoneme-related distributions. In this way, we explicitly regularize speech enhancement features by recognition vectors and the semantic information effectively helps to improve the recognizability and coherence of the enhanced speech. Experiments on public datasets demonstrate the effectiveness of PbDr in achieving both better perceptual quality and recognizability.

Chairs:

Timo Gerkmann

Tags:

signal processing society

IEEE icassp 2021

virtual conference

2021

sps

virtual conference icassp 2021

june 6-11 2021

icassp 2021

Phoneme-Based Distribution Regularization For Speech Enhancement

Yajing Liu, Xiulian Peng, Zhiwei Xiong, Yan Lu

Value-Added Bundle(s) Including this Product

ICASSP 2021 Virtual Conference - Presentation Videos Product Bundle

More Like This

Keynote: Navigating the Transition to Sustainable Energy Solutions in a Power-Hungry World

Panel: Leveraging Technology to Achieve Carbon Neutrality of Buildings and Factories

Panel: Charting the Course for Future-Ready Data Centers in the Era of Sustainability

Join the IEEE Signal Processing Society