Unsupervised Feature Learning for Environmental Sound Classification Using Cycle Consistent Generative Adversarial Network

Esmaeilpour, Mohammad; Cardinal, Patrick; Koerich, Alessandro L.

Computer Science > Machine Learning

arXiv:1904.04221v1 (cs)

[Submitted on 8 Apr 2019 (this version), latest version 25 Nov 2019 (v2)]

Title:Unsupervised Feature Learning for Environmental Sound Classification Using Cycle Consistent Generative Adversarial Network

Authors:Mohammad Esmaeilpour, Patrick Cardinal, Alessandro L. Koerich

View PDF

Abstract:In this paper we propose a novel environmental sound classification approach incorporating unsupervised feature learning from codebook via spherical $K$-Means++ algorithm and a new architecture for high-level data augmentation. The audio signal is transformed into a 2D representation using a discrete wavelet transform (DWT). The DWT spectrograms are then augmented by a novel architecture for cycle-consistent generative adversarial network. This high-level augmentation bootstraps generated spectrograms in both intra and inter class manners by translating structural features from sample to sample. A codebook is built by coding the DWT spectrograms with the speeded-up robust feature detector (SURF) and the K-Means++ algorithm. The Random Forest is our final learning algorithm which learns the environmental sound classification task from the clustered codewords in the codebook. Experimental results in four benchmarking environmental sound datasets (ESC-10, ESC-50, UrbanSound8k, and DCASE-2017) have shown that the proposed classification approach outperforms the state-of-the-art classifiers in the scope, including advanced and dense convolutional neural networks such as AlexNet and GoogLeNet, improving the classification rate between 3.51% and 14.34%, depending on the dataset.

Subjects:	Machine Learning (cs.LG); Sound (cs.SD); Audio and Speech Processing (eess.AS); Machine Learning (stat.ML)
Cite as:	arXiv:1904.04221 [cs.LG]
	(or arXiv:1904.04221v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1904.04221

Submission history

From: Mohammad Esmaeilpour [view email]
[v1] Mon, 8 Apr 2019 17:44:14 UTC (2,052 KB)
[v2] Mon, 25 Nov 2019 17:43:32 UTC (2,062 KB)

Computer Science > Machine Learning

Title:Unsupervised Feature Learning for Environmental Sound Classification Using Cycle Consistent Generative Adversarial Network

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Unsupervised Feature Learning for Environmental Sound Classification Using Cycle Consistent Generative Adversarial Network

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators