QCRI/Prop2Hate-Meme
Prop2Hate-Meme This repository presents the first Arabic Prop2Hate-Meme dataset which explore the intersection of propaganda and hate in memes using a multi-agent LLM-based framework. We extend an existing propagandistic meme dataset by annotating it with fine- and coarse-grained hate speech labels, and provide baseline experiments to support future research. Table of contents: Dataset Licensing Citation Dataset We adopted the ArMeme dataset for both fine-… See the full description on the dataset page: https://huggingface.co/datasets/QCRI/Prop2Hate-Meme.
Prop2Hate-Meme
This repository presents the first Arabic Prop2Hate-Meme dataset which explore the intersection of propaganda and hate in memes using a multi-agent LLM-based framework. We extend an existing propagandistic meme dataset by annotating it with fine- and coarse-grained hate speech labels, and provide baseline experiments to support future research.

Table of contents:
Dataset
We adopted the ArMeme dataset for both fine- and coarse-grained hatefulness categorization. We preserved the original train, development, and test splits. While ArMeme was initially annotated with four labels, for this study we retained only the memes labeled as propaganda and not_propaganda. These were subsequently re-annotated with hatefulness categories. The data distribution is provided below.
📊 Dataset Statistics
🏋️♂️ Train Split
`prop_label`
propaganda: 603not_propaganda: 1540
`hate_label`
not-hateful: 1930hateful: 213
`hate_fine_grained_label`
sarcasm: 105humor: 1815inciting violence: 13mocking: 133other: 10exclusion: 6dehumanizing: 12contempt: 38inferiority: 4slurs: 7
🧪 Dev Split
`prop_label`
not_propaganda: 224propaganda: 88
`hate_label`
not-hateful: 281hateful: 31
`hate_fine_grained_label`
humor: 260sarcasm: 19mocking: 19contempt: 7other: 1dehumanizing: 2inferiority: 1slurs: 1inciting violence: 2
🧾 Dev-Test Split (`dev_test`)
`prop_label`
not_propaganda: 436propaganda: 170
`hate_label`
not-hateful: 452hateful: 154
`hate_fine_grained_label`
humor: 334sarcasm: 118inciting violence: 12slurs: 29other: 20mocking: 49contempt: 25inferiority: 14dehumanizing: 2exclusion: 3
Experimental Scripts
Please find the experimental scripts here: https://github.com/firojalam/propaganda-and-hateful-memes.git
Licensing
This dataset is licensed under CC BY-NC-SA 4.0. To view a copy of this license, visit https://creativecommons.org/licenses/by-nc-sa/4.0/
Citation
If you use our dataset in a scientific publication, we would appreciate using the following citations:

@inproceedings{alam2024propaganda,
title={Propaganda to Hate: A Multimodal Analysis of Arabic Memes with Multi-agent LLMs},
author={Alam, Firoj and Biswas, Md Rafiul and Shah, Uzair and Zaghouani, Wajdi and Mikros, Georgios},
booktitle={International Conference on Web Information Systems Engineering},
pages={380--390},
year={2024},
organization={Springer}
}
@inproceedings{alam2024armeme,
title={{ArMeme}: Propagandistic Content in Arabic Memes},
author={Alam, Firoj and Hasnat, Abul and Ahmed, Fatema and Hasan, Md Arid and Hasanain, Maram},
booktitle={Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing (EMNLP)},
year={2024},
address={Miami, Florida},
month={November 12--16},
publisher={Association for Computational Linguistics},
}