Is it Fake? News Disinformation Detection on South African News Websites

de Wet, Harm; Marivate, Vukosi

Computer Science > Computation and Language

arXiv:2108.02941 (cs)

[Submitted on 6 Aug 2021 (v1), last revised 9 Aug 2021 (this version, v2)]

Title:Is it Fake? News Disinformation Detection on South African News Websites

Authors:Harm de Wet, Vukosi Marivate

View PDF

Abstract:Disinformation through fake news is an ongoing problem in our society and has become easily spread through social media. The most cost and time effective way to filter these large amounts of data is to use a combination of human and technical interventions to identify it. From a technical perspective, Natural Language Processing (NLP) is widely used in detecting fake news. Social media companies use NLP techniques to identify the fake news and warn their users, but fake news may still slip through undetected. It is especially a problem in more localised contexts (outside the United States of America). How do we adjust fake news detection systems to work better for local contexts such as in South Africa. In this work we investigate fake news detection on South African websites. We curate a dataset of South African fake news and then train detection models. We contrast this with using widely available fake news datasets (from mostly USA website). We also explore making the datasets more diverse by combining them and observe the differences in behaviour in writing between nations' fake news using interpretable machine learning.

Comments:	6 pages, Accepted and to be published in AFRICON 2021
Subjects:	Computation and Language (cs.CL); Computers and Society (cs.CY); Machine Learning (cs.LG)
Cite as:	arXiv:2108.02941 [cs.CL]
	(or arXiv:2108.02941v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2108.02941

Submission history

From: Vukosi Marivate [view email]
[v1] Fri, 6 Aug 2021 04:54:03 UTC (3,129 KB)
[v2] Mon, 9 Aug 2021 17:23:05 UTC (3,130 KB)

Computer Science > Computation and Language

Title:Is it Fake? News Disinformation Detection on South African News Websites

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Is it Fake? News Disinformation Detection on South African News Websites

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators