Abstract
This paper proposes an active learning method to control a labeling process for efficient annotation of acoustic training material, which is used for training sound event classifiers. The proposed method performs K-medoids clustering over an initially unlabeled dataset, and medoids as local representatives, are presented to an annotator for manual annotation. The annotated label on a medoid propagates to other samples in its cluster for label prediction. After annotating the medoids, the annotation continues to the unexamined sounds with mismatched prediction results from two classifiers, a nearest-neighbor classifier and a model-based classifier, both trained with annotated data. The annotation on the segments with mismatched predictions are ordered by the distance to the nearest annotated sample, farthest first. The evaluation is made on a public environmental sound dataset. The labels obtained through a labeling process controlled by the proposed method are used to train a classifier, using supervised learning. Only 20% of the data needs to be manually annotated with the proposed method, to achieve the accuracy with all the data annotated. In addition, the proposed method clearly outperforms other active learning algorithms proposed for sound event classification through all the experiments, simulating varying fraction of data that is manually labeled.
Original language | English |
---|---|
Title of host publication | 16th International Workshop on Acoustic Signal Enhancement, IWAENC 2018 |
Publisher | IEEE |
Pages | 116-120 |
Number of pages | 5 |
ISBN (Electronic) | 9781538681510 |
DOIs | |
Publication status | Published - 2 Nov 2018 |
Publication type | A4 Article in a conference publication |
Event | International Workshop on Acoustic Signal Enhancement - Tokyo, Japan Duration: 17 Sep 2018 → 20 Sep 2018 |
Conference
Conference | International Workshop on Acoustic Signal Enhancement |
---|---|
Country/Territory | Japan |
City | Tokyo |
Period | 17/09/18 → 20/09/18 |
Keywords
- Active learning
- Committee-based sample selection
- K-medoids clustering
- Sound event classification
Publication forum classification
- Publication forum level 1
ASJC Scopus subject areas
- Signal Processing
- Acoustics and Ultrasonics