tbogaers/RAID
RAID: Referential Availability in Implied Discourse Dataset Summary The RAID (Referential Availability in Implied Discourse) dataset is a custom synthetic dataset designed to evaluate if Large Language Models (LLMs) track the dynamic availability of discourse entities through implied cues. The dataset is meant for probing and behavioural evaluation to see if an LLM has the ability to maintain an internal situation model of referential availability, whether an… See the full description on the dataset page: https://huggingface.co/datasets/tbogaers/RAID.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face