Detection-Fusion for Knowledge Graph Extraction from Videos

Das, Taniya; Mahon, Louis; Lukasiewicz, Thomas

Computer Science > Computer Vision and Pattern Recognition

arXiv:2501.00136 (cs)

[Submitted on 30 Dec 2024]

Title:Detection-Fusion for Knowledge Graph Extraction from Videos

Authors:Taniya Das, Louis Mahon, Thomas Lukasiewicz

View PDF HTML (experimental)

Abstract:One of the challenging tasks in the field of video understanding is extracting semantic content from video inputs. Most existing systems use language models to describe videos in natural language sentences, but this has several major shortcomings. Such systems can rely too heavily on the language model component and base their output on statistical regularities in natural language text rather than on the visual contents of the video. Additionally, natural language annotations cannot be readily processed by a computer, are difficult to evaluate with performance metrics and cannot be easily translated into a different natural language. In this paper, we propose a method to annotate videos with knowledge graphs, and so avoid these problems. Specifically, we propose a deep-learning-based model for this task that first predicts pairs of individuals and then the relations between them. Additionally, we propose an extension of our model for the inclusion of background knowledge in the construction of knowledge graphs.

Comments:	12 pages, To be submitted to a conference
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2501.00136 [cs.CV]
	(or arXiv:2501.00136v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2501.00136

Submission history

From: Taniya Das [view email]
[v1] Mon, 30 Dec 2024 20:26:11 UTC (3,602 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Detection-Fusion for Knowledge Graph Extraction from Videos

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Detection-Fusion for Knowledge Graph Extraction from Videos

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators