Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation

Murugesan, Balamurali; Hussain, Rukhshanda; Bhattacharya, Rajarshi; Ayed, Ismail Ben; Dolz, Jose

Computer Science > Computer Vision and Pattern Recognition

arXiv:2307.00097v2 (cs)

[Submitted on 30 Jun 2023 (v1), revised 8 Jul 2023 (this version, v2), latest version 13 Jan 2024 (v3)]

Title:Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation

Authors:Balamurali Murugesan, Rukhshanda Hussain, Rajarshi Bhattacharya, Ismail Ben Ayed, Jose Dolz

View PDF

Abstract:Recently, CLIP-based approaches have exhibited remarkable performance on generalization and few-shot learning tasks, fueled by the power of contrastive language-vision pre-training. In particular, prompt tuning has emerged as an effective strategy to adapt the pre-trained language-vision models to downstream tasks by employing task-related textual tokens. Motivated by this progress, in this work we question whether other fundamental problems, such as weakly supervised semantic segmentation (WSSS), can benefit from prompt tuning. Our findings reveal two interesting observations that shed light on the impact of prompt tuning on WSSS. First, modifying only the class token of the text prompt results in a greater impact on the Class Activation Map (CAM), compared to arguably more complex strategies that optimize the context. And second, the class token associated with the image ground truth does not necessarily correspond to the category that yields the best CAM. Motivated by these observations, we introduce a novel approach based on a PrOmpt cLass lEarning (POLE) strategy. Through extensive experiments we demonstrate that our simple, yet efficient approach achieves SOTA performance in a well-known WSSS benchmark. These results highlight not only the benefits of language-vision models in WSSS but also the potential of prompt learning for this problem. The code is available at this https URL.

Comments:	Under review
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2307.00097 [cs.CV]
	(or arXiv:2307.00097v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2307.00097

Submission history

From: Balamurali Murugesan [view email]
[v1] Fri, 30 Jun 2023 19:25:18 UTC (6,124 KB)
[v2] Sat, 8 Jul 2023 15:51:51 UTC (6,124 KB)
[v3] Sat, 13 Jan 2024 18:23:07 UTC (6,124 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators