Learning Explicit and Implicit Arabic Discourse Relations. - Archive ouverte HAL Accéder directement au contenu
Article Dans Une Revue Journal of King Saud University - Science Année : 2014

Learning Explicit and Implicit Arabic Discourse Relations.

Résumé

We propose in this paper a supervised learning approach to identify discourse relations in Arabic texts. To our knowledge, this work represents the first attempt to focus on both explicit and implicit relations that link adjacent as well as non adjacent Elementary Discourse Units (EDUs) within the Segmented Discourse Representation Theory (SDRT). We use the Discourse Arabic Treebank corpus (D-ATB) which is composed of newspaper documents extracted from the syntactically annotated Arabic Treebank v3.2 part3 where each document is associated with complete discourse graph according to the cognitive principles of SDRT. Our list of discourse relations is composed of a three-level hierarchy of 24 relations grouped into 4 top-level classes. To automatically learn them, we use state of the art features whose efficiency has been empirically proved. We investigate how each feature contributes to the learning process. We report our experiments on identifying fine-grained discourse relations, mid-level classes and also top-level classes. We compare our approach with three baselines that are based on the most frequent relation, discourse connectives and the features used by Al-Saif and Markert (2011). Our results are very encouraging and outperform all the baselines with an F-score of 78.1% and an accuracy of 80.6%.
Fichier principal
Vignette du fichier
KESKES_16828.pdf (33.86 Mo) Télécharger le fichier
Origine : Fichiers éditeurs autorisés sur une archive ouverte

Dates et versions

hal-03519651 , version 1 (10-01-2022)

Identifiants

Citer

Iskandar Keskes, Farah Benamara, Lamia Hadrich Belguith. Learning Explicit and Implicit Arabic Discourse Relations.. Journal of King Saud University - Science, 2014, 26 (4), pp.398-416. ⟨10.1016/j.jksuci.2014.06.001⟩. ⟨hal-03519651⟩
13 Consultations
3 Téléchargements

Altmetric

Partager

Gmail Facebook X LinkedIn More