Harnessing Multilingual Resources to Question Answering in Arabic

Alnajjar, Khalid; Hämäläinen, Mika

Computer Science > Computation and Language

arXiv:2205.08024 (cs)

[Submitted on 16 May 2022]

Title:Harnessing Multilingual Resources to Question Answering in Arabic

Authors:Khalid Alnajjar, Mika Hämäläinen

View PDF

Abstract:The goal of the paper is to predict answers to questions given a passage of Qur'an. The answers are always found in the passage, so the task of the model is to predict where an answer starts and where it ends. As the initial data set is rather small for training, we make use of multilingual BERT so that we can augment the training data by using data available for languages other than Arabic. Furthermore, we crawl a large Arabic corpus that is domain specific to religious discourse. Our approach consists of two steps, first we train a BERT model to predict a set of possible answers in a passage. Finally, we use another BERT based model to rank the candidate answers produced by the first BERT model.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2205.08024 [cs.CL]
	(or arXiv:2205.08024v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2205.08024

Submission history

From: Mika Hämäläinen [view email]
[v1] Mon, 16 May 2022 23:28:01 UTC (967 KB)

Full-text links:

Access Paper:

view license

Current browse context:

< prev | next >

new | recent | 2022-05

Change to browse by:

cs.CL

References & Citations

export BibTeX citation

Computer Science > Computation and Language

Title:Harnessing Multilingual Resources to Question Answering in Arabic

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Harnessing Multilingual Resources to Question Answering in Arabic

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators