상세 보기
Performance analysis of weakly-supervised sound event detection system based on the mean-teacher convolutional recurrent neural network model
WEB OF SCIENCE
0SCOPUS
0초록
This paper introduces and implements a Sound Event Detection (SED) system based on weakly-supervised learning where only part of the data is labeled, and analyzes the effect of parameters. The SED system estimates the classes and onset/offset times of events in the acoustic signal. In order to train the model, all information on the event class and onset/offset times must be provided. Unfortunately, the onset/offset times are hard to be labeled exactly. Therefore, in the weakly-supervised task, the SED model is trained by "strongly labeled data" including the event class and activations, "weakly labeled data" including the event class, and "unlabeled data" without any label. Recently, the SED systems using the mean-teacher model are widely used for the task with several parameters. These parameters should be chosen carefully because they may affect the performance. In this paper, performance analysis was performed on parameters, such as the feature, moving average parameter, weight of the consistency cost function, ramp-up length, and maximum learning rate, using the data of DCASE 2020 Task 4. Effects and the optimal values of the parameters were discussed.
키워드
- 제목
- Performance analysis of weakly-supervised sound event detection system based on the mean-teacher convolutional recurrent neural network model
- 저자
- Lee, Seokjin
- 발행일
- 2021
- 유형
- Article
- 저널명
- 한국음향학회지
- 권
- 40
- 호
- 2
- 페이지
- 139 ~ 147
- 언어
- KOR
- 출판사
- ACOUSTICAL SOC KOREA
- 발행국가
- 대한민국
- 분량
- 9 페이지
- ISSN
- E 2287-3775
P 1225-4428