IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks

Huang, Zitong; Chen, Ze; Dong, Bowen; Liang, Chaoqi; Zhou, Erjin; Zuo, Wangmeng

Computer Science > Computer Vision and Pattern Recognition

arXiv:2404.16331 (cs)

[Submitted on 25 Apr 2024 (v1), last revised 4 Dec 2024 (this version, v2)]

Title:IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks

Authors:Zitong Huang, Ze Chen, Bowen Dong, Chaoqi Liang, Erjin Zhou, Wangmeng Zuo

View PDF HTML (experimental)

Abstract:Model Weight Averaging (MWA) is a technique that seeks to enhance model's performance by averaging the weights of multiple trained models. This paper first empirically finds that 1) the vanilla MWA can benefit the class-imbalanced learning, and 2) performing model averaging in the early epochs of training yields a greater performance improvement than doing that in later epochs. Inspired by these two observations, in this paper we propose a novel MWA technique for class-imbalanced learning tasks named Iterative Model Weight Averaging (IMWA). Specifically, IMWA divides the entire training stage into multiple episodes. Within each episode, multiple models are concurrently trained from the same initialized model weight, and subsequently averaged into a singular model. Then, the weight of this average model serves as a fresh initialization for the ensuing episode, thus establishing an iterative learning paradigm. Compared to vanilla MWA, IMWA achieves higher performance improvements with the same computational cost. Moreover, IMWA can further enhance the performance of those methods employing EMA strategy, demonstrating that IMWA and EMA can complement each other. Extensive experiments on various class-imbalanced learning tasks, i.e., class-imbalanced image classification, semi-supervised class-imbalanced image classification and semi-supervised object detection tasks showcase the effectiveness of our IMWA.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2404.16331 [cs.CV]
	(or arXiv:2404.16331v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2404.16331

Submission history

From: Zitong Huang [view email]
[v1] Thu, 25 Apr 2024 04:37:35 UTC (3,317 KB)
[v2] Wed, 4 Dec 2024 07:47:10 UTC (3,307 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators