flashcards
medical_meadow_medical_flashcards
Dataset Card for Medical Flashcards
Dataset Summary
Medicine as a whole encompasses a wide range of subjects that medical students and graduates must master
in order to practice effectively. This includes a deep understanding of basic medical sciences, clinical knowledge,
and clinical skills. The Anki Medical Curriculum flashcards are created and updated by medical students and cover the
entirety of this curriculum, addressing subjects such as anatomy, physiology… See the full description on the dataset page: https://huggingface.co/datasets/medalpaca/medical_meadow_medical_flashcards.medical-meadow-medical-flashcards
Dataset Card for medical-meadow-medical-flashcards
This dataset originates from the medAlpaca repository.
The medical-meadow-medical-flashcards dataset is specifically used for models training of medical question-answering.
Dataset Details
Dataset Description
Each sample is comprised of three columns: instruction, input and output.
Language(s): English
Dataset Sources
The code from the original repository was adopted to post it here.
Repository:… See the full description on the dataset page: https://huggingface.co/datasets/flwrlabs/medical-meadow-medical-flashcards.dolma-reddit-to-flashcards-0625
Overview
Dolma Reddit to Flashcards is a dataset of synthetically-generated QA items created on the basis of filtered Reddit data.
The creation of this dataset was motivated by the observation in Dolma (Soldaini et al. 2024) that the original Dolma Reddit data showed no benefit from inclusion of thread-level context over isolated submissions and comments, and that clean performance distinctions between tested Reddit versions were limited mainly to the HellaSWAG benchmark.
The… See the full description on the dataset page: https://huggingface.co/datasets/allenai/dolma-reddit-to-flashcards-0625.medical_meadow_medical_flashcardsmedical-meadow-medical-flashcards-splits
Dataset Card for Medical Meadow Medical Flashcards - Fixed Splits
Dataset Summary
This dataset is a reproducible train/validation/test split of
flwrlabs/medical-meadow-medical-flashcards,
an English medical question-answering dataset from the
MedAlpaca project.
The source dataset contains 33,955 flashcards in a single training split. This
version preserves the original rows and columns while assigning every example
to one of three fixed splits using seed 42. No… See the full description on the dataset page: https://huggingface.co/datasets/leandrodevai/medical-meadow-medical-flashcards-splits.buzz_sources_049_medical_meadow_medical_flashcards
