{# Audit 04/10/2026 : « autre » n'est pas un code de langue ; SPHAERO n'est pas l'éditeur des documents qu'elle héberge ou référence. #} {# citation_pdf_url doit mener à un PDF : un lien vers une page DOI est pénalisé par Google Scholar (avant : tout lien externe). #}
Accès ouvert · CC BY

Offensive Language Detection in Arabizi

Article scientifique 2023 Anglais

Résumé

Detecting offensive language in underresourced languages presents a significant real-world challenge for social media platforms.This paper is the first work focused on the issue of offensive language detection in Arabizi, an under-explored topic in an under-resourced form of Arabic.For the first time, a comprehensive and critical overview of the existing work on the topic is presented.In addition, we carry out experiments using different BERT-like models and show the feasibility of detecting offensive language in Arabizi with high accuracy.Throughout a thorough analysis of results, we emphasize the complexities introduced by dialect variations and out-ofdomain generalization.We use in our experiments a dataset that we have constructed by leveraging existing, albeit limited, resources.To facilitate further research, we make this dataset publicly accessible to the research community.

Citer ce document

Bensalem, I., Mout, M. L., & Rosso, P. (2023). Offensive Language Detection in Arabizi. https://doi.org/10.18653/v1/2023.arabicnlp-1.36

Exporter : BibTeX · RIS (Zotero, Mendeley, EndNote)

Accès au document

Texte intégral en lecture en ligne, réservé aux abonnés SPHAERO et aux membres de l'institution. Se connecter

Voir l'article sur le site de la revue

Licence et provenance

Licence : CC BY

Notice moissonnée depuis OpenAlex le 10/09/2026. Le document reste hébergé par sa source.
Voir le document à la source →

Statistiques

Consultations : 2

Téléchargements : 0