Compressed Video Action Recognition with Refined Motion Vector

Cao, Haoyuan; Yu, Shining; Feng, Jiashi

Electrical Engineering and Systems Science > Image and Video Processing

arXiv:1910.02533 (eess)

[Submitted on 6 Oct 2019]

Title:Compressed Video Action Recognition with Refined Motion Vector

Authors:Haoyuan Cao, Shining Yu, Jiashi Feng

View PDF

Abstract:Although CNN has reached satisfactory performance in image-related tasks, using CNN to process videos is much more challenging due to the enormous size of raw video streams. In this work, we propose to use motion vectors and residuals from modern video compression techniques to effectively learn the representation of the raw frames and greatly remove the temporal redundancy, giving a faster video processing model. Compressed Video Action Recognition(CoViAR) has explored to directly use compressed video to train the deep neural network, where the motion vectors were utilized to present temporal information. However, motion vector is designed for minimizing video size where precise motion information is not obligatory. Compared with optical flow, motion vectors contain noisy and unreliable motion information. Inspired by the mechanism of video compression codecs, we propose an approach to refine the motion vectors where unreliable movement will be removed while temporal information is largely reserved. We prove that replacing the original motion vector with refined one and using the same network as CoViAR has achieved state-of-art performance on the UCF-101 and HMDB-51 with negligible efficiency degrades comparing with original CoViAR.

Comments:	8 pages, 3 figures, 4 tables
Subjects:	Image and Video Processing (eess.IV)
Cite as:	arXiv:1910.02533 [eess.IV]
	(or arXiv:1910.02533v1 [eess.IV] for this version)
	https://doi.org/10.48550/arXiv.1910.02533

Submission history

From: Haoyuan Cao [view email]
[v1] Sun, 6 Oct 2019 21:34:42 UTC (1,856 KB)

Electrical Engineering and Systems Science > Image and Video Processing

Title:Compressed Video Action Recognition with Refined Motion Vector

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Image and Video Processing

Title:Compressed Video Action Recognition with Refined Motion Vector

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators