CoolFace
Datasetpublicgated

elliot-mllm/ScienceQA_RS_think

ScienceQA — ScienceQA_RS_think Rejection-sampled from the ScienceQA train split. This split holds the accepted items, with the model's reasoning trace. rows 4,630 QA pairs 13,834 shards 12 accepted / rejected (whole family) 13,834 / 641 accept rate 95.6% verifier anls How the data was produced A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with the official ground truth by the verifier… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/ScienceQA_RS_think.

sourceHugging Faceotherupdated 26d agoView on Hugging Face
0likes42downloads

No commit history came back for main. The revision may not exist, or the source declined the request.