CoolFace
Datasetpublic

gaffarshaikh07/rag-qa-evaluation-dataset

RAG QA Evaluation Dataset Overview This dataset contains test cases for evaluating Retrieval-Augmented Generation (RAG) and Large Language Model (LLM) applications. The dataset is designed from a software testing and quality engineering perspective. Dataset Structure Each test case contains: Field Description question User question sent to the AI application context Context available to the AI application expected_answer Expected… See the full description on the dataset page: https://huggingface.co/datasets/gaffarshaikh07/rag-qa-evaluation-dataset.

sourceHugging Facemitupdated 10d agoView on Hugging Face
1likes67downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
gaffarshaikh07/rag-qa-evaluation-dataset · CoolFace