elliot-mllm/WordArt_RS_think
WordArt — WordArt_RS_think Rejection-sampled from the WordArt train split. This split holds the accepted items, with the model's reasoning trace. rows 1,846 QA pairs 1,846 shards 35 accepted / rejected (whole family) 1,846 / 2,958 accept rate 38.4% verifier exact How the data was produced A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with the official ground truth by the verifier described… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/WordArt_RS_think.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face