CoolFace
Datasetpublic

RKB109/rag-evaluation-lab-20260908-dataset

RAG Evaluation Lab Synthetic Dataset Summary This dataset contains 14 training examples and 4 held-out examples for RAG systems often ship without a stable regression set or failure taxonomy. Every record is synthetic and includes: input: query, event, or feature description label: expected class, route, relation, or evidence category context: synthetic supporting context source: fictional source identifier variant: generation pattern synthetic: always true… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/rag-evaluation-lab-20260908-dataset.

sourceHugging Facecc-by-4.0updated 15d agoView on Hugging Face
0likes58downloads
Dataset Card

RAG Evaluation Lab Synthetic Dataset

Summary

This dataset contains 14 training examples and 4 held-out examples for RAG systems often ship without a stable regression set or failure taxonomy.

Every record is synthetic and includes:

  • input: query, event, or feature description
  • label: expected class, route, relation, or evidence category
  • context: synthetic supporting context
  • source: fictional source identifier
  • variant: generation pattern
  • synthetic: always true

Uses

  • Reproducible unit and integration tests
  • Baseline model training
  • Evaluation harness development
  • Schema and architecture demonstrations

Limitations

Synthetic cases validate the harness, not a production RAG system. Teams must add representative domain examples.

This dataset does not represent real users, patients, customers, production traffic, or licensed media. It must not be presented as real-world evidence.

Related Model

RKB109/rag-evaluation-lab-20260908-model