Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Contrastive Environmental Sound Representation Learning

Jul 18, 2022

Peter Ochieng, Dennis Kaburu

Figure 1 for Contrastive Environmental Sound Representation Learning

Figure 2 for Contrastive Environmental Sound Representation Learning

Figure 3 for Contrastive Environmental Sound Representation Learning

Figure 4 for Contrastive Environmental Sound Representation Learning

Share this with someone who'll enjoy it:

Abstract:Machine hearing of the environmental sound is one of the important issues in the audio recognition domain. It gives the machine the ability to discriminate between the different input sounds that guides its decision making. In this work we exploit the self-supervised contrastive technique and a shallow 1D CNN to extract the distinctive audio features (audio representations) without using any explicit annotations.We generate representations of a given audio using both its raw audio waveform and spectrogram and evaluate if the proposed learner is agnostic to the type of audio input. We further use canonical correlation analysis (CCA) to fuse representations from the two types of input of a given audio and demonstrate that the fused global feature results in robust representation of the audio signal as compared to the individual representations. The evaluation of the proposed technique is done on both ESC-50 and UrbanSound8K. The results show that the proposed technique is able to extract most features of the environmental audio and gives an improvement of 12.8% and 0.9% on the ESC-50 and UrbanSound8K datasets respectively.

View paper on

Share this with someone who'll enjoy it:

Title:Contrastive Environmental Sound Representation Learning

Paper and Code