JOINT NOISE REDUCTION AND LISTENING ENHANCEMENT FOR FULL-END SPEECH ENHANCEMENT

Haoyu Li (National Institute of Informatics); Yun Liu (National Institute of Informatics); Junichi Yamagishi (National Institute of Informatics)

DOI

SPS

Members: Free
IEEE Members: $11.00
Non-members: $15.00

07 Jun 2023

Speech enhancement (SE) methods mainly focus on recovering clean speech from noisy input. In real-world speech communication, however, noises often exist in not only speaker but also listener environments. Although SE methods can suppress the noise contained in the speaker’s voice, they cannot deal with the noise that is physically present in the listener side. To address such a complicated but common scenario, we investigate a deep learning-based joint framework integrating noise reduction (NR) with listening enhancement (LE), in which the NR module first suppresses noise and the LE module then modifies the denoised speech, i.e., the output of the NR module, to further improve speech intelligibility. The enhanced speech can thus be less noisy and more intelligible for listeners. Experimental results show that our proposed method achieves promising results and significantly outperforms the disjoint processing methods in terms of various speech evaluation metrics.

Tags:

Audio signal enhancement and restoration

JOINT NOISE REDUCTION AND LISTENING ENHANCEMENT FOR FULL-END SPEECH ENHANCEMENT

Haoyu Li (National Institute of Informatics); Yun Liu (National Institute of Informatics); Junichi Yamagishi (National Institute of Informatics)

Value-Added Bundle(s) Including this Product

IEEE ICASSP 2023, 4-10 June 2023, Greece. Virtual and In-Person Conference - Presentation Videos Product Bundle

More Like This

MAID: A Conditional Diffusion Model For Long Music Audio Inpainting

Immersive enhancement and removal of loudspeaker sound using wireless assistive listening systems and binaural hearing devices

Extreme Audio Time Stretching using Neural Synthesis

Join the IEEE Signal Processing Society