CoolFace
Datasetpublic

ctoraman/large-scale-hate-speech-turkish-v2

The dataset published in the LREC 2022 paper "Large-Scale Hate Speech Detection with Cross-Domain Transfer". This is Dataset v2 (Turkish): The modified dataset that includes 60,310 tweets in Turkish. The annotations with more than 80% agreement are included. TweetID: Tweet ID from Twitter API LangID: 0 (Turkish) TopicID: Domain of the topic 0-Religion, 1-Gender, 2-Race, 3-Politics, 4-Sports HateLabel: Final hate label decision 0-Normal, 1-Offensive, 2-Hate… See the full description on the dataset page: https://huggingface.co/datasets/ctoraman/large-scale-hate-speech-turkish-v2.

sourceHugging Facecc-by-nc-sa-4.0updated 2y agoView on Hugging Face
0likes40downloads

ctoraman/large-scale-hate-speech-turkish-v2 · main · files are served by the source, never re-hosted here