LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool

Ma, Yue; Ren, Huantao; Wang, Boyu; Jin, Jingang; Velipasalar, Senem; Qiu, Qinru

Computer Science > Computer Vision and Pattern Recognition

arXiv:2412.05840 (cs)

[Submitted on 8 Dec 2024]

Title:LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool

Authors:Yue Ma, Huantao Ren, Boyu Wang, Jingang Jin, Senem Velipasalar, Qinru Qiu

View PDF HTML (experimental)

Abstract:Continual learning aims to update a model so that it can sequentially learn new tasks without forgetting previously acquired knowledge. Recent continual learning approaches often leverage the vision-language model CLIP for its high-dimensional feature space and cross-modality feature matching. Traditional CLIP-based classification methods identify the most similar text label for a test image by comparing their embeddings. However, these methods are sensitive to the quality of text phrases and less effective for classes lacking meaningful text labels. In this work, we rethink CLIP-based continual learning and introduce the concept of Label Vector Pool (LVP). LVP replaces text labels with training images as similarity references, eliminating the need for ideal text descriptions. We present three variations of LVP and evaluate their performance on class and domain incremental learning tasks. Leveraging CLIP's high dimensional feature space, LVP learning algorithms are task-order invariant. The new knowledge does not modify the old knowledge, hence, there is minimum forgetting. Different tasks can be learned independently and in parallel with low computational and memory demands. Experimental results show that proposed LVP-based methods outperform the current state-of-the-art baseline by a significant margin of 40.7%.

Comments:	submitted to CVPR2025
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
MSC classes:	68T45
ACM classes:	I.2.10; I.4; I.5
Cite as:	arXiv:2412.05840 [cs.CV]
	(or arXiv:2412.05840v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2412.05840

Submission history

From: Yue Ma [view email]
[v1] Sun, 8 Dec 2024 07:22:39 UTC (39,815 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators