An Evolutionary-based Generative Approach for Audio Data Augmentation

Silvan Mertes, Alice Baird, Dominik Schiller, Björn Schuller, Elisabeth André

DOI

SPS

Members: Free
IEEE Members: $11.00
Non-members: $15.00

Length: 04:45

21 Sep 2020

In this paper, we introduce a novel framework to augment raw audio data for machine learning classification tasks. For the first part of our framework, we employ a generative adversarial network (GAN) to create new variants of the audio samples that are already existing in our source dataset for the classification task. In the second step, we then utilize an evolutionary algorithm to search the input domain space of the previously trained GAN, with respect to predefined characteristics of the generated audio. This way we are able to generate audio in a controlled manner that contributes to an improvement in classification performance of the original task. To validate our approach, we chose to test it on the task of soundscape classification. We show that our approach leads to a substantial improvement in classification results when compared to a training routine without data augmentation and training with uncontrolled data augmentation with GANs.

Tags:

sps conference

virtual workshop

mmsp 2020

September 2020

An Evolutionary-based Generative Approach for Audio Data Augmentation

Silvan Mertes, Alice Baird, Dominik Schiller, Björn Schuller, Elisabeth André

Value-Added Bundle(s) Including this Product

MMSP 2020 Virtual Conference - Presentation Videos Product Bundle

More Like This

IEEE ICASSP 2023, 4-10 June 2023, Greece. Virtual and In-Person Conference - Presentation Videos Product Bundle

IEEE ICASSP 2024, 1 4-19 April 2024, Seoul, Korea. Conference Presentation Videos Bundle

ICIP 2022, October 16-19, 2022, Bordeaux, France - Presentation Videos Product Bundle

Join the IEEE Signal Processing Society