Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:A Computational Acquisition Model for Multimodal Word Categorization

May 12, 2022

Uri Berger, Gabriel Stanovsky, Omri Abend, Lea Frermann

Figure 1 for A Computational Acquisition Model for Multimodal Word Categorization

Figure 2 for A Computational Acquisition Model for Multimodal Word Categorization

Figure 3 for A Computational Acquisition Model for Multimodal Word Categorization

Figure 4 for A Computational Acquisition Model for Multimodal Word Categorization

Share this with someone who'll enjoy it:

Abstract:Recent advances in self-supervised modeling of text and images open new opportunities for computational models of child language acquisition, which is believed to rely heavily on cross-modal signals. However, prior studies have been limited by their reliance on vision models trained on large image datasets annotated with a pre-defined set of depicted object categories. This is (a) not faithful to the information children receive and (b) prohibits the evaluation of such models with respect to category learning tasks, due to the pre-imposed category structure. We address this gap, and present a cognitively-inspired, multimodal acquisition model, trained from image-caption pairs on naturalistic data using cross-modal self-supervision. We show that the model learns word categories and object recognition abilities, and presents trends reminiscent of those reported in the developmental literature. We make our code and trained models public for future reference and use.

* Accepted to NAACL 2022

View paper on

Share this with someone who'll enjoy it:

Title:A Computational Acquisition Model for Multimodal Word Categorization

Paper and Code