todor-cmd/memdoc
memdoc A dataset for evaluating QA agents that retrieve from both conversational memory and a document corpus, while controlling for variance that comes from the question and its evidence themselves. Each item is a question with 2-4 gold-evidence chunks partitioned across the two stores. The same questions and evidence can be presented as memory-only, document-only, or split across both. Because the question text and the underlying facts stay fixed, differences in agent… See the full description on the dataset page: https://huggingface.co/datasets/todor-cmd/memdoc.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face