Knowledge Graph Enhanced Generative Multi-modal Models for Class-Incremental Learning

Cao, Xusheng; Lu, Haori; Huang, Linlan; Yang, Fei; Liu, Xialei; Cheng, Ming-Ming

Computer Science > Computer Vision and Pattern Recognition

arXiv:2503.18403 (cs)

[Submitted on 24 Mar 2025]

Title:Knowledge Graph Enhanced Generative Multi-modal Models for Class-Incremental Learning

Authors:Xusheng Cao, Haori Lu, Linlan Huang, Fei Yang, Xialei Liu, Ming-Ming Cheng

View PDF HTML (experimental)

Abstract:Continual learning in computer vision faces the critical challenge of catastrophic forgetting, where models struggle to retain prior knowledge while adapting to new tasks. Although recent studies have attempted to leverage the generalization capabilities of pre-trained models to mitigate overfitting on current tasks, models still tend to forget details of previously learned categories as tasks progress, leading to misclassification. To address these limitations, we introduce a novel Knowledge Graph Enhanced Generative Multi-modal model (KG-GMM) that builds an evolving knowledge graph throughout the learning process. Our approach utilizes relationships within the knowledge graph to augment the class labels and assigns different relations to similar categories to enhance model differentiation. During testing, we propose a Knowledge Graph Augmented Inference method that locates specific categories by analyzing relationships within the generated text, thereby reducing the loss of detailed information about old classes when learning new knowledge and alleviating forgetting. Experiments demonstrate that our method effectively leverages relational information to help the model correct mispredictions, achieving state-of-the-art results in both conventional CIL and few-shot CIL settings, confirming the efficacy of knowledge graphs at preserving knowledge in the continual learning scenarios.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2503.18403 [cs.CV]
	(or arXiv:2503.18403v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2503.18403

Submission history

From: Xialei Liu [view email]
[v1] Mon, 24 Mar 2025 07:20:43 UTC (8,939 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Knowledge Graph Enhanced Generative Multi-modal Models for Class-Incremental Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Knowledge Graph Enhanced Generative Multi-modal Models for Class-Incremental Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators