elliot-mllm/OrandCar_rejected
OrandCar — OrandCar_rejected Rejection-sampled from the OrandCar train split. This split holds the rejected items — the answer field holds the official ground truth. rows 445 QA pairs 445 shards 1 accepted / rejected (whole family) 1,576 / 445 accept rate 78.0% verifier alnum The rejected split is training data, not just diagnostics: answer is the official ground truth, and wrong_vlm records what the model said instead. How the data was… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/OrandCar_rejected.
024
No card is published for this repository, or it could not be fetched from Hugging Face right now.
