humor-detection
ColBERT_Humor_Detection
ColBERT_Humor
Dataset Summary
ColBERT Humor contains 200,000 labeled short texts, equally distributed between humorous and non-humorous content. The dataset was created to overcome the limitations of prior humor detection datasets, which were characterized by inconsistencies in text length, word count, and formality, making them easy to predict with simple models without truly understanding the nuances of humor. The two sources for this dataset are the News Category… See the full description on the dataset page: https://huggingface.co/datasets/CreativeLang/ColBERT_Humor_Detection.encoded_humor_detection_3gn-humor-detection
Text-based afective computing
We collected a dataset of tweets primarily written in Guarani (and Jopara, a code-switching language that combines Guarani and Spanish) and annotated them for three widely-used dimensions in sentiment analysis:
emotion recognition (https://huggingface.co/datasets/mmaguero/gn-emotion-recognition),
humor detection (this repo, https://huggingface.co/datasets/mmaguero/gn-humor-detection), and
offensive language identification… See the full description on the dataset page: https://huggingface.co/datasets/mmaguero/gn-humor-detection.encoded_humor_detection_2encoded_humor_detection_4encoded_humor_detection_1
