Accès ouvert

A Character Level Convolutional BiLSTM for Arabic Dialect Identification

Article scientifique 2019 Anglais

Résumé

In this paper, we describe the contribution of CU-RAISA team to the 2019 Madar shared task 2 1 , which focused on Twitter User finegrained dialect identification. Among participating teams, our system ranked the 4th (with 61.54%) F1-Macro measure. Our system is trained using a character level convolutional bidirectional long-short-term memory (BiL-STM) network trained on approximately 2k users' data. We show that training on concatenated user tweets as input is further superior to training on user tweets separately and assign user's label on the mode of user's tweets' predictions.

Citer ce document

Elaraby, M., Zahran, A. (2019). A Character Level Convolutional BiLSTM for Arabic Dialect Identification. https://doi.org/10.18653/v1/w19-4636

Accès au document

Voir sur le dépôt source

Ce document est hébergé sur son dépôt institutionnel d'origine.

Statistiques

Consultations : 1

Téléchargements : 0