datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ragas-wikiqa
Dataset Card for "ragas-wikiqa"
More Information needed
EnDev_RAGAS_testset
EnDev RAGAS Test Set
Synthetic Q&A test set (42 pairs) generated with RAGAS
(TestsetGenerator.generate_with_chunks) over chunks sampled from the EnDev corpus
stored in Qdrant collection endev-bgem3-512 (Gradio-gateway Space
GIZ/EnDev_Qdrant).
Generated: 2026-09-14 UTC
Generator/judge LLM: Qwen/Qwen3-235B-A22B-Instruct-2507
Embeddings: BGE-M3 via the EnDev TEI Inference Endpoint
Columns: user_input, reference, reference_contexts, synthesizer_name
Used to evaluate the deployed… See the full description on the dataset page: https://huggingface.co/datasets/GIZ/EnDev_RAGAS_testset.ragas-golden-dataset
Dataset Card for the ragas-golden-dataset
Dataset Description
The RAGAS Golden Dataset is a synthetically generated question-answering dataset designed for evaluating Retrieval Augmented Generation (RAG) systems. It contains high-quality question-answer pairs derived from academic papers on AI agents and agentic AI architectures.
Dataset Summary
This dataset was generated using Prefect and the RAGAS TestsetGenerator framework, which creates synthetic questions… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/ragas-golden-dataset.RAGAS_xquad_x_squadtest_half split from XQuAD https://huggingface.co/datasets/xxizhouu/RAGAS_xquad
PLUS
one impossible question(english) for each paragraph, taken from SQuAD 2.0
test_id: shared uuid accross different spilt
cmi: code mix index
rag_assets_10292024_filtered_v1ragas-webgpt
Dataset Card for "ragas-webgpt"
More Information needed
ragas-golden-dataset-documents
Dataset Card for RAGAS Golden Dataset Documents
A small, mixed‐format corpus to compare PDF, API, and web‐based document loader output from the LangChain ecosystem.
The code to run the Prefect prefect_docloader_pipeline.py pipeline is available in the RAGAS Golden Dataset Pipeline repository.
While several enhancements are planned for future iterations, the hands-on insights gained from this grassroots exploration of document loader behaviors proved too valuable -- things that… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/ragas-golden-dataset-documents.ragas_trainscarnatic-ragasragas-golden-dataset-v2
Dataset Card for the ragas-golden-dataset-v2
Dataset Description
The RAGAS Golden Dataset is a synthetically generated question-answering dataset designed for evaluating Retrieval Augmented Generation (RAG) systems. It contains high-quality question-answer pairs derived from academic papers on AI agents and agentic AI architectures.
Dataset Summary
This dataset was generated using Prefect and the RAGAS TestsetGenerator framework, which creates synthetic… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/ragas-golden-dataset-v2.ragas-eval-datasetRAGAS_xquaddata source: https://github.com/google-deepmind/xquad (base on squad 1.0)
document_en/de.zip contains all 48 documents for answer the question: created by concatenate continous pragraphs
Question,Answer, Context (QAC) pair in english(en) AND german(de)
clean_full_en_de: 48 documents, 5 paragraphs per document, multiple questions per paragraph
single_qa_en_de: 48 documents, 5 paragraphs per document, one question per paragraph
test_half: 24 documents, 5 paragraphs per document, one question… See the full description on the dataset page: https://huggingface.co/datasets/xxizhouu/RAGAS_xquad.ai-arxiv2-ragas-mixtralragas-test-datasetRAGAS1ragas_SSqdrant_docs_qna_ragasragas-golden-testset-personas
Dataset Card for ragas-golden-testset-personas
Dataset Description
The RAGAS Golden Dataset is a synthetically generated question-answering dataset designed for evaluating Retrieval Augmented Generation (RAG) systems. It contains high-quality question-answer pairs derived from academic papers on AI agents and agentic AI architectures.
Dataset Summary
This dataset was generated using the RAGAS TestsetGenerator framework, which creates synthetic questions… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/ragas-golden-testset-personas.ragas-airline-datasettest_ragascontext-precision-ragasragas_gardian_evaluation_overlapping
📚 GARDIAN-RAGAS QA Dataset
A synthetic question–answer (QA) dataset generated from the GARDIAN corpus using RAGAS and the open-weight Mistral-7B-Instruct-v0.3 model. This dataset is designed to support evaluation and benchmarking of retrieval-augmented generation (RAG) systems, with an emphasis on grounded, high-fidelity QA generation.
📦 Dataset Summary
Source Corpus: GARDIAN scientific article collection
QA Generation Model: Mistral-7B-Instruct-v0.3
Sample Size: 1… See the full description on the dataset page: https://huggingface.co/datasets/CGIAR/ragas_gardian_evaluation_overlapping.rag-evaluation-ragasds
Dataset Card for the ragas-golden-dataset-v2
Dataset Description
The RAGAS Golden Dataset is a synthetically generated question-answering dataset designed for evaluating Retrieval Augmented Generation (RAG) systems. It contains high-quality question-answer pairs derived from academic papers on AI agents and agentic AI architectures.
Dataset Summary
This dataset was generated using Prefect and the RAGAS TestsetGenerator framework, which creates synthetic… See the full description on the dataset page: https://huggingface.co/datasets/Gabriel10-10/rag-evaluation-ragasds.meta-record-ragas-synthetic-dataset
Dataset Card for "meta-record-ragas-synthetic-dataset"
More Information needed
ragastest_ragasragas-golden-dataset-colabtest_ragas_llamaragas-retreival_top1
Dataset Card for "ragas-retreival_top1"
More Information needed
testset_ragas
