elliot-mllm/OrandCar_RS_nothink
OrandCar — OrandCar_RS_nothink Rejection-sampled from the OrandCar train split. This split holds the accepted items, answer only. rows 1,576 QA pairs 1,576 shards 1 accepted / rejected (whole family) 1,576 / 445 accept rate 78.0% verifier alnum How the data was produced A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with the official ground truth by the verifier described below; matches go… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/OrandCar_RS_nothink.
024
No card is published for this repository, or it could not be fetched from Hugging Face right now.
