CoolFace
Datasetpublic

alan9622/marabert2-levantine-toxic-model-v4-dataset

L-HSAB (Custom Split for Levantine Hate Speech) This dataset contains Levantine Arabic tweets labeled for Hate Speech and Abusive language. It is used to train the model: amitca71/marabert2-levantine-toxic-model-v4 Dataset Structure text: The tweet content (Arabic). label: The classification label. 0: Abusive 1: Normal 2: Hate Citation Mulki, H., et al. (2019). "L-HSAB: A Levantine Twitter Dataset for Hate Speech and Abusive Language."

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes14downloads
Dataset Card

L-HSAB (Custom Split for Levantine Hate Speech)

This dataset contains Levantine Arabic tweets labeled for Hate Speech and Abusive language. It is used to train the model: amitca71/marabert2-levantine-toxic-model-v4

Dataset Structure

  • —text: The tweet content (Arabic).
  • —label: The classification label.
  • —0: Abusive
  • —1: Normal
  • —2: Hate

Citation

Mulki, H., et al. (2019). "L-HSAB: A Levantine Twitter Dataset for Hate Speech and Abusive Language."