Texture-Based Input Feature Selection for Action Recognition

Jiang, Yalong

Computer Science > Computer Vision and Pattern Recognition

arXiv:2303.00138 (cs)

[Submitted on 28 Feb 2023 (v1), last revised 23 Apr 2023 (this version, v3)]

Title:Texture-Based Input Feature Selection for Action Recognition

Authors:Yalong Jiang

View PDF

Abstract:The performance of video action recognition has been significantly boosted by using motion representations within a two-stream Convolutional Neural Network (CNN) architecture. However, there are a few challenging problems in action recognition in real scenarios, e.g., the variations in viewpoints and poses, and the changes in backgrounds. The domain discrepancy between the training data and the test data causes the performance drop. To improve the model robustness, we propose a novel method to determine the task-irrelevant content in inputs which increases the domain discrepancy. The method is based on a human parsing model (HP model) which jointly conducts dense correspondence labelling and semantic part segmentation. The predictions from the HP model also function as re-rendering the human regions in each video using the same set of textures to make humans appearances in all classes be the same. A revised dataset is generated for training and testing and makes the action recognition model exhibit invariance to the irrelevant content in the inputs. Moreover, the predictions from the HP model are used to enrich the inputs to the AR model during both training and testing. Experimental results show that our proposed model is superior to existing models for action recognition on the HMDB-51 dataset and the Penn Action dataset.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2303.00138 [cs.CV]
	(or arXiv:2303.00138v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2303.00138

Submission history

From: Allen Jiang [view email]
[v1] Tue, 28 Feb 2023 23:56:31 UTC (12,913 KB)
[v2] Sat, 4 Mar 2023 14:12:25 UTC (12,910 KB)
[v3] Sun, 23 Apr 2023 09:00:53 UTC (1,487 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Texture-Based Input Feature Selection for Action Recognition

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Texture-Based Input Feature Selection for Action Recognition

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators