There is a newer version of the record available.

Published May 9, 2023 | Version 1.0

ESC50Mix dataset

Authors/Creators

  • 1. Fraunhofer IDMT

Description

Reference:

Jakob Abeßer, Sascha Grollmisch, Meinard Müller, How Robust are Audio Embeddings for Polyphonic Sound Event Tagging? (to be published in  IEEE/ACM Transactions on Audio, Speech, and Language Processing)

The dataset builds upon the ESC-50 dataset (https://github.com/karolpiczak/ESC-50) and creates mixtures of two random sound pairs by systematically blending from one to the other sounds. 

In total, the dataset includes 30,000 5s audio clips

- 1000 sound pairs

- 6 gain factors (used for mixing pairs)

- 5 data augmentations (none, loudness reduction, Gaussian noise, high-frequency boost, low-frequency boost)

Files

ESC50MixBlend.zip

Files (10.2 GB)

Name Size
md5:7fb38b7eae63c09fd6a3fb87d66e7b81
10.2 GB Preview Download