CoolFace
Datasetpublic

ctoraman/large-scale-hate-speech-turkish-v2

The dataset published in the LREC 2022 paper "Large-Scale Hate Speech Detection with Cross-Domain Transfer". This is Dataset v2 (Turkish): The modified dataset that includes 60,310 tweets in Turkish. The annotations with more than 80% agreement are included. TweetID: Tweet ID from Twitter API LangID: 0 (Turkish) TopicID: Domain of the topic 0-Religion, 1-Gender, 2-Race, 3-Politics, 4-Sports HateLabel: Final hate label decision 0-Normal, 1-Offensive, 2-Hate… See the full description on the dataset page: https://huggingface.co/datasets/ctoraman/large-scale-hate-speech-turkish-v2.

sourceHugging Facecc-by-nc-sa-4.0updated 2y agoView on Hugging Face
0likes40downloads
5 commits on main
156a45b2y ago

Update README.md

ctoraman
83d11473y ago

Update README.md

ctoraman
c37f49c3y ago

Update README.md

ctoraman
eef08083y ago

Upload toraman22_hate_speech_dataset_tr_v2.tsv

ctoraman
1f364773y ago

initial commit

ctoraman