Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models

Mar 28, 2024

Saurav Jha, Dong Gong, Lina Yao

Figure 1 for CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models

Figure 2 for CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models

Figure 3 for CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models

Figure 4 for CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models

Share this with someone who'll enjoy it:

Abstract:Continual learning (CL) aims to help deep neural networks to learn new knowledge while retaining what has been learned. Recently, pre-trained vision-language models such as CLIP, with powerful generalization ability, have been gaining traction as practical CL candidates. However, the domain mismatch between the pre-training and the downstream CL tasks calls for finetuning of the CLIP on the latter. The deterministic nature of the existing finetuning methods makes them overlook the many possible interactions across the modalities and deems them unsafe for high-risk CL tasks requiring reliable uncertainty estimation. To address these, our work proposes Continual LeArning with Probabilistic finetuning (CLAP). CLAP develops probabilistic modeling over task-specific modules with visual-guided text features, providing more reliable fine-tuning in CL. It further alleviates forgetting by exploiting the rich pre-trained knowledge of CLIP for weight initialization and distribution regularization of task-specific modules. Cooperating with the diverse range of existing prompting methods, CLAP can surpass the predominant deterministic finetuning approaches for CL with CLIP. Lastly, we study the superior uncertainty estimation abilities of CLAP for novel data detection and exemplar selection within CL setups. Our code is available at \url{https://github.com/srvCodes/clap4clip}.

* Work under review

View paper on

Share this with someone who'll enjoy it:

Title:CLAP4CLIP: Continual Learning with Probabilistic Finetuning for Vision-Language Models

Paper and Code