datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vimqa-generated-answers-pass1
Vi-MQA - Pass 1 Generated Answers & Evaluation
This repo contains the Pass 1 outputs and evaluation results for the Vi-MQA Dataset from the VMLU Benchmark Suite with a total of 4,762 records.
Folder Structure
1. Model Outputs (raw_outputs/)
Contains the formatted outputs from the 3 models evaluated in Pass 1:
results_pass1_gemma.jsonl (Gemma 4 31B IT)
results_pass1_llama.jsonl (Llama 4 Scout)
results_pass1_qwen.jsonl (Qwen3 32B)
2.… See the full description on the dataset page: https://huggingface.co/datasets/nygdon/vimqa-generated-answers-pass1.3dx-users-guide-generated-errors-sample
