CoolFace
Datasetpublic

benchmark-data/hallucination-traps

Hallucination Traps A curated benchmark dataset consisting of intentionally misleading prompts designed to evaluate hallucination behavior in language models. Each prompt appears plausible at first glance but contains a subtle false premise, nonexistent entity, or incorrect factual assumption. The expected behavior is that the model should either refuse, express uncertainty, or explicitly identify the incorrect premise rather than hallucinate a confident but false answer.… See the full description on the dataset page: https://huggingface.co/datasets/benchmark-data/hallucination-traps.

sourceHugging Facecc-by-4.0updated 9mo agoView on Hugging Face
0likes15downloads
3 commits on main
8adbf769mo ago

Upload hallucination_traps.csv

benchmark-data
dc9aab69mo ago

Update README.md

benchmark-data
f8e1fcd9mo ago

initial commit

benchmark-data