Skip to main content

DEEPSPACE: DYNAMIC SPATIAL AND SOURCE CUE BASED SOURCE SEPARATION FOR DIALOG ENHANCEMENT

Aaron S Master (Dolby Laboratories, Inc); Lie Lu (Dolby Laboratories); Jonas Samuelsson (Dolby Laboratories, Inc); Heidi-Maria Lehtonen (Dolby Sweden AB); Scott Norcross (Dolby Laboratories, Inc); Nathan Swedlow (Dolby Laboratories, Inc); Audrey Howard (Dolby Laboratories, Inc)

  • SPS
    Members: Free
    IEEE Members: $11.00
    Non-members: $15.00
06 Jun 2023

Dialog Enhancement (DE) is a feature which allows a user to increase the level of dialog in TV or movie content relative to non-dialog sounds. When only the original mix is available, DE is “unguided,” and requires source separation. In this paper, we describe the DeepSpace system, which performs source separation using both dynamic spatial cues and source cues to support unguided DE. Its technologies include spatio-level filtering (SLF) and deep-learning based dialog classification and denoising. Using subjective listening tests, we show that DeepSpace demonstrates significantly improved overall performance relative to state-of-the-art systems available for testing. We explore the feasibility of using existing automated metrics to evaluate unguided DE systems.

More Like This

  • SPS
    Members: Free
    IEEE Members: $11.00
    Non-members: $15.00
  • SPS
    Members: Free
    IEEE Members: $11.00
    Non-members: $15.00
  • SPS
    Members: Free
    IEEE Members: $11.00
    Non-members: $15.00