gaffarshaikh07/rag-qa-evaluation-dataset
RAG QA Evaluation Dataset Overview This dataset contains test cases for evaluating Retrieval-Augmented Generation (RAG) and Large Language Model (LLM) applications. The dataset is designed from a software testing and quality engineering perspective. Dataset Structure Each test case contains: Field Description question User question sent to the AI application context Context available to the AI application expected_answer Expected… See the full description on the dataset page: https://huggingface.co/datasets/gaffarshaikh07/rag-qa-evaluation-dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face