CoolFace
Datasetpublic

disi-unibo-nlp/Phunny

Phunny: A Humor-Based QA Benchmark for Evaluating LLM Generalization Welcome to Phunny, a humor-based question answering (QA) benchmark designed to evaluate the reasoning and generalization abilities of large language models (LLMs) through structured puns. This repository accompanies our ACL 2025 main track paper:"What do you call a dog that is incontrovertibly true? Dogma: Testing LLM Generalization through Humor" To reproduce our experiments: Code available on GitHub… See the full description on the dataset page: https://huggingface.co/datasets/disi-unibo-nlp/Phunny.

sourceHugging Facemitupdated 1y agoView on Hugging Face
3likes32downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
disi-unibo-nlp/Phunny · CoolFace