CoolFace
Datasetpublic

kilizi/FactGuard

FactGuard-Bench FactGuard-Bench is a bilingual long-context benchmark for evaluating and improving whether language models answer only when the supplied document contains sufficient evidence. It contains English and Chinese examples from the book and legal domains, with contexts extending to approximately 128K in the legacy character-based construction buckets. The benchmark accompanies: Towards Reliable Long-Context Reasoning: Detecting Unanswerable Questions via FactGuard… See the full description on the dataset page: https://huggingface.co/datasets/kilizi/FactGuard.

sourceHugging Facecc-by-4.0updated 24d agoView on Hugging Face
0likes93downloads
settings

This repository belongs to kilizi on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameFactGuard
visibilitypublic
licencecc-by-4.0
gatedno
ownerkilizi
Account settings
kilizi/FactGuard · CoolFace