elliot-mllm/ScienceQA_RS_think
ScienceQA — ScienceQA_RS_think Rejection-sampled from the ScienceQA train split. This split holds the accepted items, with the model's reasoning trace. rows 4,630 QA pairs 13,834 shards 12 accepted / rejected (whole family) 13,834 / 641 accept rate 95.6% verifier anls How the data was produced A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with the official ground truth by the verifier… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/ScienceQA_RS_think.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face