CoolFace
Datasetpublic

OctoReasoner/mercury_verl

Mercury (verl efficiency eval set) The eval split of Elfsong/Mercury (arXiv 2402.07844; 256 LeetCode-style tasks; the train split ships no test cases and is not gradable), converted to the verl rule-reward schema by verl/scripts/data/mercury.py. Source license CC-BY-NC-4.0 (non-commercial) -- this conversion keeps that license. Every row's ground truth carries the full official scoring contract: entry point, the task's convert_offline/evaluate_offline hooks (lctk linked-list /… See the full description on the dataset page: https://huggingface.co/datasets/OctoReasoner/mercury_verl.

sourceHugging Facecc-by-nc-4.0updated 1mo agoView on Hugging Face
0likes75downloads
5 commits on main
0cb63251mo ago

dataset card: provenance, scoring contract, compromises

wetsoledrysoul
a12e0f71mo ago

Mercury eval split in verl schema (official-harness scoring, unified training-format prompts)

wetsoledrysoul
1d884cb2mo ago

dataset card: provenance, scoring contract, compromises

wetsoledrysoul
d7ba9252mo ago

Mercury eval split in verl schema (official-harness scoring contract)

wetsoledrysoul
2e755b42mo ago

initial commit

wetsoledrysoul