HyperAIHyperAI

Command Palette

Search for a command to run...

3 months ago

Heavily Augmented Sound Event Detection utilizing Weak Predictions

Hyeonuk Nam Byeong-Yun Ko Gyeong-Tae Lee Seong-Hu Kim Won-Ho Jung Sang-Min Choi Yong-Hwa Park

Heavily Augmented Sound Event Detection utilizing Weak Predictions

Abstract

The performances of Sound Event Detection (SED) systems are greatly limited by the difficulty in generating large strongly labeled dataset. In this work, we used two main approaches to overcome the lack of strongly labeled data. First, we applied heavy data augmentation on input features. Data augmentation methods used include not only conventional methods used in speech/audio domains but also our proposed method named FilterAugment. Second, we propose two methods to utilize weak predictions to enhance weakly supervised SED performance. As a result, we obtained the best PSDS1 of 0.4336 and best PSDS2 of 0.8161 on the DESED real validation dataset. This work is submitted to DCASE 2021 Task4 and is ranked on the 3rd place. Code availa-ble: https://github.com/frednam93/FilterAugSED.

Code Repositories

frednam93/FilterAugSED
Official
pytorch
Mentioned in GitHub

Benchmarks

BenchmarkMethodologyMetrics
sound-event-detection-on-desedFiltAug SED
PSDS1: 0.4336
PSDS2: 0.8161
event-based F1 score: 49.6

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp
Heavily Augmented Sound Event Detection utilizing Weak Predictions | Papers | HyperAI