datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
COVID-QA-unique-context-test-10-percent-validation-10-percent
Dataset Card for "COVID-QA-unique-context-test-10-percent-validation-10-percent"
More Information needed
downstream_validation_qatruthful_qa-validation-german_q_n_aqa_validation_qwenCOVID-QA-train-80-test-10-validation-10
Dataset Card for "COVID-QA-train-80-test-10-validation-10"
More Information needed
vietnamese_legal_closed_qa_validation
Evaluation dataset for Vietnamse Legal RAG Chatbot
The dataset is manually filterred and refined by humans. Each sample in the dataset contains a question, K contexts and an answer generated by OpenAI GPT-4. The answer is refined by humans.
COVID-QA-validation-sentence-transformer
Dataset Card for "COVID-QA-validation-sentence-transformer"
More Information needed
hotpotqa-distractor-qa-with-ids-validationhotpotqa-distractor-qa-with-ids-validationsquad-qa-validationThis dataset processed version of the SQuAD dataset, which is provided by Hugging Face. The SQuAD dataset is a collection of questions and answers derived from a set of Wikipedia articles, designed for machine reading comprehension tasks.
Dataset source: https://huggingface.co/datasets/rajpurkar/squad
