datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
offensive-humor@article{tang2022naughtyformer,
title={The Naughtyformer: A Transformer Understands Offensive Humor},
author={Tang, Leonard and Cai, Alexander and Li, Steve and Wang, Jason},
journal={arXiv preprint arXiv:2211.14369},
year={2022}
}
ColBERT_Humor_Detection
ColBERT_Humor
Dataset Summary
ColBERT Humor contains 200,000 labeled short texts, equally distributed between humorous and non-humorous content. The dataset was created to overcome the limitations of prior humor detection datasets, which were characterized by inconsistencies in text length, word count, and formality, making them easy to predict with simple models without truly understanding the nuances of humor. The two sources for this dataset are the News Category… See the full description on the dataset page: https://huggingface.co/datasets/CreativeLang/ColBERT_Humor_Detection.Barcenas-HumorNegroDataset en español con 500 chistes de humor negro y una explicación.
Datos creados de manera sintética por Claude 3 Haiku y Llama 3 70B Instruct.
El proceso para crear el dataset fue el recopilar de varias fuentes chistes de humor negro en español para luego ser utilizadas en los mejores modelos como Gemini 1.5 Pro, Claude 3, etc.
Con eso genere cientos de chistes de humor negro en español para tener más datos y hacer un super recopilatorio de chistes de humor negro en español, aproximadamente… See the full description on the dataset page: https://huggingface.co/datasets/Danielbrdz/Barcenas-HumorNegro.KoWit-24
KoWit-24
Slides | Prompts
Overview
We present KoWit-24, a dataset with fine-grained annotation of wordplay in 2,700 Russian news headlines. KoWit-24 annotations include the presence of wordplay, its type, wordplay anchors, and words/phrases the wordplay refers to.
Content
Overview
ContentDataset
Description
Download
Key features
How to load and use
Experiments
Wordplay detection
Wordplay interpretation
Automatic interpretation evaluation
Table… See the full description on the dataset page: https://huggingface.co/datasets/Humor-Research/KoWit-24.slpg_humor_generationHumor_Votehumor_trainannotations_creators: []
language_creators: []
languages: []
licenses: []
multilinguality: []
pretty_name: humor_train
size_categories: []
source_datasets: []
task_categories: []
task_ids: []
HumorDatasetSmallHUMORDatasetLargeHUMORDatasetMediumGreek-Humor-DatasetThe Greek Humor Dataset (GHD). The first ever balanced and human-annotated Humor Dataset in Greek language.
