datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PMC-VQA-text
PMC-VQA-text
This dataset is a text format of PMC-VQA.
We built this dataset using the Meta-Llama-3-70B-Instruct, and the instruction we used is: Rewrite the question-answer pairs into a paragraph format (Do not use the words 'question' and 'answer' in your responses):.
train_text.json corresponds to the train.csv and train_2.csv splits in the PMC-VQA dataset.
Samples with two or more question-and-answer pairs were selected.
Citation
If you find this dataset useful… See the full description on the dataset page: https://huggingface.co/datasets/myeongkyunkang/PMC-VQA-text.PMC-VQA
