bhumika-tewari-282006/halluciguard-benchmark
HalluciGuard Benchmark A 500-sample benchmark for evaluating hallucination detection in Retrieval-Augmented Generation (RAG) systems, built for the paper HalluciGuard: A Label-Free Confidence-Aware Hallucination Detection Framework for Retrieval-Augmented Generation Systems. Construction Each sample is built from a hand-verified atomic fact rather than downloaded from an existing QA corpus. Samples are constructed to mirror the query phrasing, difficulty, and… See the full description on the dataset page: https://huggingface.co/datasets/bhumika-tewari-282006/halluciguard-benchmark.
Upload README.md with huggingface_hub
Upload README.md with huggingface_hub
Upload benchmark_data.json with huggingface_hub
initial commit
