CLMay 12, 2022

DTW at Qur'an QA 2022: Utilising Transfer Learning with Transformers for Question Answering in a Low-resource Domain

arXiv:2205.06025v1586 citationsh-index: 34
Originality Synthesis-oriented
AI Analysis

This addresses the understudied problem of machine reading comprehension for religious texts, specifically the Qur'an, but is incremental as it applies existing methods to a new domain.

The paper tackled question answering on the Qur'an, a low-resource domain, by using transfer learning with transformers and ensemble strategies, achieving a partial Reciprocal Rank score of 0.49 on the test set.

The task of machine reading comprehension (MRC) is a useful benchmark to evaluate the natural language understanding of machines. It has gained popularity in the natural language processing (NLP) field mainly due to the large number of datasets released for many languages. However, the research in MRC has been understudied in several domains, including religious texts. The goal of the Qur'an QA 2022 shared task is to fill this gap by producing state-of-the-art question answering and reading comprehension research on Qur'an. This paper describes the DTW entry to the Quran QA 2022 shared task. Our methodology uses transfer learning to take advantage of available Arabic MRC data. We further improve the results using various ensemble learning strategies. Our approach provided a partial Reciprocal Rank (pRR) score of 0.49 on the test set, proving its strong performance on the task.

Code Implementations1 repo
Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes