Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Sri Lakshmi

Contrastive Language-Image Pre-training for the Italian Language

Aug 19, 2021

Federico Bianchi, Giuseppe Attanasio, Raphael Pisoni, Silvia Terragni, Gabriele Sarti, Sri Lakshmi

Figure 1 for Contrastive Language-Image Pre-training for the Italian Language

Figure 2 for Contrastive Language-Image Pre-training for the Italian Language

Figure 3 for Contrastive Language-Image Pre-training for the Italian Language

Figure 4 for Contrastive Language-Image Pre-training for the Italian Language

Abstract:CLIP (Contrastive Language-Image Pre-training) is a very recent multi-modal model that jointly learns representations of images and texts. The model is trained on a massive amount of English data and shows impressive performance on zero-shot classification tasks. Training the same model on a different language is not trivial, since data in other languages might be not enough and the model needs high-quality translations of the texts to guarantee a good performance. In this paper, we present the first CLIP model for the Italian Language (CLIP-Italian), trained on more than 1.4 million image-text pairs. Results show that CLIP-Italian outperforms the multilingual CLIP model on the tasks of image retrieval and zero-shot classification.

Via

Access Paper or Ask Questions