Localization-Aware Multi-Scale Representation Learning for Repetitive Action Counting

Wang, Sujia; Shen, Xiangwei; Tang, Yansong; Dong, Xin; Geng, Wenjia; Chen, Lei

Abstract:Repetitive action counting (RAC) aims to estimate the number of class-agnostic action occurrences in a video without exemplars. Most current RAC methods rely on a raw frame-to-frame similarity representation for period prediction. However, this approach can be significantly disrupted by common noise such as action interruptions and inconsistencies, leading to sub-optimal counting performance in realistic scenarios. In this paper, we introduce a foreground localization optimization objective into similarity representation learning to obtain more robust and efficient video features. We propose a Localization-Aware Multi-Scale Representation Learning (LMRL) framework. Specifically, we apply a Multi-Scale Period-Aware Representation (MPR) with a scale-specific design to accommodate various action frequencies and learn more flexible temporal correlations. Furthermore, we introduce the Repetition Foreground Localization (RFL) method, which enhances the representation by coarsely identifying periodic actions and incorporating global semantic information. These two modules can be jointly optimized, resulting in a more discerning periodic action representation. Our approach significantly reduces the impact of noise, thereby improving counting accuracy. Additionally, the framework is designed to be scalable and adaptable to different types of video content. Experimental results on the RepCountA and UCFRep datasets demonstrate that our proposed method effectively handles repetitive action counting.

Comments:	Accepted by IEEE VCIP2024
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2501.07312 [cs.CV]
	(or arXiv:2501.07312v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2501.07312

Computer Science > Computer Vision and Pattern Recognition

Title:Localization-Aware Multi-Scale Representation Learning for Repetitive Action Counting

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators