Foundation Model Drives Weakly Incremental Learning for Semantic Segmentation

Yu, Chaohui; Zhou, Qiang; Li, Jingliang; Yuan, Jianlong; Wang, Zhibin; Wang, Fan

Computer Science > Computer Vision and Pattern Recognition

arXiv:2302.14250 (cs)

[Submitted on 28 Feb 2023 (v1), last revised 20 Apr 2023 (this version, v2)]

Title:Foundation Model Drives Weakly Incremental Learning for Semantic Segmentation

Authors:Chaohui Yu, Qiang Zhou, Jingliang Li, Jianlong Yuan, Zhibin Wang, Fan Wang

View PDF

Abstract:Modern incremental learning for semantic segmentation methods usually learn new categories based on dense annotations. Although achieve promising results, pixel-by-pixel labeling is costly and time-consuming. Weakly incremental learning for semantic segmentation (WILSS) is a novel and attractive task, which aims at learning to segment new classes from cheap and widely available image-level labels. Despite the comparable results, the image-level labels can not provide details to locate each segment, which limits the performance of WILSS. This inspires us to think how to improve and effectively utilize the supervision of new classes given image-level labels while avoiding forgetting old ones. In this work, we propose a novel and data-efficient framework for WILSS, named FMWISS. Specifically, we propose pre-training based co-segmentation to distill the knowledge of complementary foundation models for generating dense pseudo labels. We further optimize the noisy pseudo masks with a teacher-student architecture, where a plug-in teacher is optimized with a proposed dense contrastive loss. Moreover, we introduce memory-based copy-paste augmentation to improve the catastrophic forgetting problem of old classes. Extensive experiments on Pascal VOC and COCO datasets demonstrate the superior performance of our framework, e.g., FMWISS achieves 70.7% and 73.3% in the 15-5 VOC setting, outperforming the state-of-the-art method by 3.4% and 6.1%, respectively.

Comments:	CVPR 2023
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2302.14250 [cs.CV]
	(or arXiv:2302.14250v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2302.14250

Submission history

From: Chaohui Yu [view email]
[v1] Tue, 28 Feb 2023 02:21:42 UTC (7,997 KB)
[v2] Thu, 20 Apr 2023 08:12:44 UTC (7,997 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Foundation Model Drives Weakly Incremental Learning for Semantic Segmentation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Foundation Model Drives Weakly Incremental Learning for Semantic Segmentation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators