Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision

Nov 14, 2019

Aren Jansen, Daniel P. W. Ellis, Shawn Hershey, R. Channing Moore, Manoj Plakal, Ashok C. Popat, Rif A. Saurous

Figure 1 for Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision

Figure 2 for Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision

Figure 3 for Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision

Figure 4 for Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision

Share this with someone who'll enjoy it:

Abstract:Humans do not acquire perceptual abilities in the way we train machines. While machine learning algorithms typically operate on large collections of randomly-chosen, explicitly-labeled examples, human acquisition relies more heavily on multimodal unsupervised learning (as infants) and active learning (as children). With this motivation, we present a learning framework for sound representation and recognition that combines (i) a self-supervised objective based on a general notion of unimodal and cross-modal coincidence, (ii) a clustering objective that reflects our need to impose categorical structure on our experiences, and (iii) a cluster-based active learning procedure that solicits targeted weak supervision to consolidate categories into relevant semantic classes. By training a combined sound embedding/clustering/classification network according to these criteria, we achieve a new state-of-the-art unsupervised audio representation and demonstrate up to a 20-fold reduction in the number of labels required to reach a desired classification performance.

* This extended version of a ICASSP 2020 submission under same title has an added figure and additional discussion for easier consumption

View paper on

Share this with someone who'll enjoy it:

Title:Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision

Paper and Code