benchmark-data/hallucination-traps
Hallucination Traps A curated benchmark dataset consisting of intentionally misleading prompts designed to evaluate hallucination behavior in language models. Each prompt appears plausible at first glance but contains a subtle false premise, nonexistent entity, or incorrect factual assumption. The expected behavior is that the model should either refuse, express uncertainty, or explicitly identify the incorrect premise rather than hallucinate a confident but false answer.… See the full description on the dataset page: https://huggingface.co/datasets/benchmark-data/hallucination-traps.
015
Upload hallucination_traps.csv
Update README.md
initial commit
