Performance analysis of weakly-supervised sound event detection system based on the mean-teacher convolutional recurrent neural network model

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

This paper introduces and implements a Sound Event Detection (SED) system based on weakly-supervised learning where only part of the data is labeled, and analyzes the effect of parameters. The SED system estimates the classes and onset/offset times of events in the acoustic signal. In order to train the model, all information on the event class and onset/offset times must be provided. Unfortunately, the onset/offset times are hard to be labeled exactly. Therefore, in the weakly-supervised task, the SED model is trained by "strongly labeled data" including the event class and activations, "weakly labeled data" including the event class, and "unlabeled data" without any label. Recently, the SED systems using the mean-teacher model are widely used for the task with several parameters. These parameters should be chosen carefully because they may affect the performance. In this paper, performance analysis was performed on parameters, such as the feature, moving average parameter, weight of the consistency cost function, ramp-up length, and maximum learning rate, using the data of DCASE 2020 Task 4. Effects and the optimal values of the parameters were discussed.

키워드

Sound event detection; Semi-supervised learning; Mean-teacher; Convolutional recurrent neural network
제목
Performance analysis of weakly-supervised sound event detection system based on the mean-teacher convolutional recurrent neural network model
저자
Lee, Seokjin
DOI
10.7776/ASK.2021.40.2.139
발행일
2021
유형
Article
저널명
한국음향학회지
권
40
호
2
페이지
139 ~ 147