RKB109/rag-evaluation-lab-20260908-dataset
RAG Evaluation Lab Synthetic Dataset Summary This dataset contains 14 training examples and 4 held-out examples for RAG systems often ship without a stable regression set or failure taxonomy. Every record is synthetic and includes: input: query, event, or feature description label: expected class, route, relation, or evidence category context: synthetic supporting context source: fictional source identifier variant: generation pattern synthetic: always true… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/rag-evaluation-lab-20260908-dataset.
RAG Evaluation Lab Synthetic Dataset
Summary
This dataset contains 14 training examples and 4 held-out examples for RAG systems often ship without a stable regression set or failure taxonomy.
Every record is synthetic and includes:
input: query, event, or feature descriptionlabel: expected class, route, relation, or evidence categorycontext: synthetic supporting contextsource: fictional source identifiervariant: generation patternsynthetic: alwaystrue
Uses
- Reproducible unit and integration tests
- Baseline model training
- Evaluation harness development
- Schema and architecture demonstrations
Limitations
Synthetic cases validate the harness, not a production RAG system. Teams must add representative domain examples.
This dataset does not represent real users, patients, customers, production traffic, or licensed media. It must not be presented as real-world evidence.
