datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
COVID-QA-unique-context-test-10-percent-validation-10-percent
Dataset Card for "COVID-QA-unique-context-test-10-percent-validation-10-percent"
More Information needed
downstream_validation_qatruthful_qa-validation-german_q_n_atrivia_qa_validation_mc
TriviaQA Shortcut
Overview
This repository contains a multiple-choice reformulation of the TriviaQA dataset benchmark, created as part of our work on reducing the computational cost of evaluating large language models (LLMs) during (pre-)training.
Each item contains an answers column, where the first item is always the original correct answer from TriviaQA. The remaining answer options are distractors generated by Meta-Llama-3.1-70B-Instruct-GPTQ-INT4. For details… See the full description on the dataset page: https://huggingface.co/datasets/fraunhofer-iis/trivia_qa_validation_mc.qa_validation_qwenCOVID-QA-train-80-test-10-validation-10
Dataset Card for "COVID-QA-train-80-test-10-validation-10"
More Information needed
vietnamese_legal_closed_qa_validation
Evaluation dataset for Vietnamse Legal RAG Chatbot
The dataset is manually filterred and refined by humans. Each sample in the dataset contains a question, K contexts and an answer generated by OpenAI GPT-4. The answer is refined by humans.
COVID-QA-validation-sentence-transformer
Dataset Card for "COVID-QA-validation-sentence-transformer"
More Information needed
hotpotqa-distractor-qa-with-ids-validationbarexam_qa_size_60_validation_modelknowledge_validation_not_grounded_barexam_qahotpotqa-distractor-qa-with-ids-validationCOVID-QA-2-unique-context-test-10-percent-validation-10-percentCOVID-QA-1-unique-context-test-10-percent-validation-10-percentsquad-qa-validationThis dataset processed version of the SQuAD dataset, which is provided by Hugging Face. The SQuAD dataset is a collection of questions and answers derived from a set of Wikipedia articles, designed for machine reading comprehension tasks.
Dataset source: https://huggingface.co/datasets/rajpurkar/squad
hotpot_qa_validation_model_size_50Medical_QA_Validationset
