amitca71/marabert2-levantine-toxic-model-v3-dataset
L-HSAB (Custom Split for Levantine Hate Speech) This dataset contains Levantine Arabic tweets labeled for Hate Speech and Abusive language. It is used to train the model: amitca71/marabert2-levantine-toxic-model-v3 Dataset Structure text: The tweet content (Arabic). label: The classification label. 0: Abusive 1: Normal 2: Hate Citation Mulki, H., et al. (2019). "L-HSAB: A Levantine Twitter Dataset for Hate Speech and Abusive Language."
019
Upload README.md with huggingface_hub
Upload dataset
initial commit
