CoolFace
Datasetpublic

DIYIN/ElephantBench

ElephantBench ElephantBench is a closed-book knowledge probe for evaluating whether a language model remembers long-tail facts and recalls the different verified accounts associated with them. The release contains 1,094 English questions. Evaluation code, prompts, construction utilities, and full documentation are available in the ElephantBench GitHub repository. Load the dataset from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/DIYIN/ElephantBench.

sourceHugging Facecc-by-4.0updated 23d agoView on Hugging Face
0likes114downloads

DIYIN/ElephantBench · main · files are served by the source, never re-hosted here