ViMGuard: A Novel Multi-Modal System for Video Misinformation Guarding

Kan, Andrew; Kan, Christopher; Nabulsi, Zaid

Abstract:The rise of social media and short-form video (SFV) has facilitated a breeding ground for misinformation. With the emergence of large language models, significant research has gone into curbing this misinformation problem with automatic false claim detection for text. Unfortunately, the automatic detection of misinformation in SFV is a more complex problem that remains largely unstudied. While text samples are monomodal (only containing words), SFVs comprise three different modalities: words, visuals, and non-linguistic audio. In this work, we introduce Video Masked Autoencoders for Misinformation Guarding (ViMGuard), the first deep-learning architecture capable of fact-checking an SFV through analysis of all three of its constituent modalities. ViMGuard leverages a dual-component system. First, Video and Audio Masked Autoencoders analyze the visual and non-linguistic audio elements of a video to discern its intention; specifically whether it intends to make an informative claim. If it is deemed that the SFV has informative intent, it is passed through our second component: a Retrieval Augmented Generation system that validates the factual accuracy of spoken words. In evaluation, ViMGuard outperformed three cutting-edge fact-checkers, thus setting a new standard for SFV fact-checking and marking a significant stride toward trustworthy news on social platforms. To promote further testing and iteration, VimGuard was deployed into a Chrome extension and all code was open-sourced on GitHub.

Comments:	7 pages, 2 figures
Subjects:	Machine Learning (cs.LG); Computation and Language (cs.CL); Computers and Society (cs.CY)
Cite as:	arXiv:2410.16592 [cs.LG]
	(or arXiv:2410.16592v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2410.16592

Computer Science > Machine Learning

Title:ViMGuard: A Novel Multi-Modal System for Video Misinformation Guarding

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators