Event Masked Autoencoder: Point-wise Action Recognition with Event-Based Cameras

Sun, Jingkai; Zhang, Qiang; Wang, Jiaxu; Cao, Jiahang; Xu, Renjing

Computer Science > Computer Vision and Pattern Recognition

arXiv:2501.01040 (cs)

[Submitted on 2 Jan 2025]

Title:Event Masked Autoencoder: Point-wise Action Recognition with Event-Based Cameras

Authors:Jingkai Sun, Qiang Zhang, Jiaxu Wang, Jiahang Cao, Renjing Xu

View PDF HTML (experimental)

Abstract:Dynamic vision sensors (DVS) are bio-inspired devices that capture visual information in the form of asynchronous events, which encode changes in pixel intensity with high temporal resolution and low latency. These events provide rich motion cues that can be exploited for various computer vision tasks, such as action recognition. However, most existing DVS-based action recognition methods lose temporal information during data transformation or suffer from noise and outliers caused by sensor imperfections or environmental factors. To address these challenges, we propose a novel framework that preserves and exploits the spatiotemporal structure of event data for action recognition. Our framework consists of two main components: 1) a point-wise event masked autoencoder (MAE) that learns a compact and discriminative representation of event patches by reconstructing them from masked raw event camera points data; 2) an improved event points patch generation algorithm that leverages an event data inlier model and point-wise data augmentation techniques to enhance the quality and diversity of event points patches. To the best of our knowledge, our approach introduces the pre-train method into event camera raw points data for the first time, and we propose a novel event points patch embedding to utilize transformer-based models on event cameras.

Comments:	ICASSP 2025 Camera Ready
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2501.01040 [cs.CV]
	(or arXiv:2501.01040v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2501.01040

Submission history

From: Jingkai Sun [view email]
[v1] Thu, 2 Jan 2025 03:49:03 UTC (1,819 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Event Masked Autoencoder: Point-wise Action Recognition with Event-Based Cameras

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Event Masked Autoencoder: Point-wise Action Recognition with Event-Based Cameras

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators