FelipeGuerra/Colombian_Spanish_Cyberbullying_Dataset_2
Dataset Summary This dataset consists of 2566 tweets and maintains a balanced distribution between cyberbullying and not cyberbullying. For every keyword or phrase, there is an annotated tweet labeled as cyberbullying that contains that word or phrase. The not cyberbullying category predominantly includes tweets that do not contain obscene words and are sourced from popular and varied discussions involving colombian users, reflecting a wide range of topics and conversations. The… See the full description on the dataset page: https://huggingface.co/datasets/FelipeGuerra/Colombian_Spanish_Cyberbullying_Dataset_2.
219
