iCompass Working Notes for the Nuanced Arabic Dialect Identification Shared task
Abir Messaoudi, Chayma Fourati, Hatem Haddad, Moez BenHajhmida · 2022
We describe our submitted system to the Nuanced Arabic Dialect Identification (NADI) shared task.We tackled only the first subtask (Subtask 1).We used state-of-the-art Deep Learning models and pre-trained contextualized text representation models that we finetuned according to the downstream task in hand.As a first approach, we used BERT Arabic variants: MARBERT with its two versions MARBERT v1 and MARBERT v2, then we combined MARBERT embeddings with a CNN classifier, and finally, we tested the Quasi-Recurrent Neural Networks (QRNN) model.The results found show that version 2 of MAR-BERT outperforms all of the previously mentioned models on Subtask 1.