CoolFace
Datasetpublic

ZuoXiaojia/medical-qa-datasets

all-processed dataset is a concatenation of of medical-meadow-* and chatdoctor_healthcaremagic datasets The Chat Doctor term is replaced by the chatbot term in the chatdoctor_healthcaremagic dataset Similar to the literature the medical_meadow_cord19 dataset is subsampled to 50,000 samples truthful-qa-* is a benchmark dataset for evaluating the truthfulness of models in text generation, which is used in Llama 2 paper. Within this dataset, there are 55 and 16 questions related to Health and… See the full description on the dataset page: https://huggingface.co/datasets/ZuoXiaojia/medical-qa-datasets.

sourceHugging Faceupdated 8mo agoView on Hugging Face
0likes196downloads
Dataset Card
  • —all-processed dataset is a concatenation of of medical-meadow-* and chatdoctor_healthcaremagic datasets
  • —The Chat Doctor term is replaced by the chatbot term in the chatdoctor_healthcaremagic dataset
  • —Similar to the literature the medical_meadow_cord19 dataset is subsampled to 50,000 samples
  • —truthful-qa-* is a benchmark dataset for evaluating the truthfulness of models in text generation, which is used in Llama 2 paper. Within this dataset, there are 55 and 16 questions related to Health and Nutrition, respectively, making it a valuable resource for medical question-answering scenarios.