datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
unified-toolcalls-canonical
Unified Tool-Calling Corpus — Canonicalized Output
Publish-ready conversion of two pinned Hugging Face dataset revisions into the single
schema defined in docs/unified_format.md, with repeated
records normalized by an explicit canonicalization rule and every surviving record
kept faithful to its source row.
Records in (source rows)
65,000
Records published (canonical survivors)
64,622
Duplicates collapsed
378 (343 duplicate groups)
Records mutated during… See the full description on the dataset page: https://huggingface.co/datasets/dongbobo/unified-toolcalls-canonical.tool-calls-single-reasoningmoh_8_rollouts_tool_calls
MOH 8 Rollouts Tool Calls
Fixed 8-rollout benchmark sampled from aimosprite/training-output using only tool-calling attempts.
Construction
Source problem family: polymath_*
Source eligibility band: correct_count_16 in [1, 10]
Candidate attempt pool per problem: attempts with Python Calls > 0
Final sample per problem: 8 attempts, sampled without replacement using seed 42
Additional constraint: the sampled 8 always include at least one correct attempt
Published rows: 300… See the full description on the dataset page: https://huggingface.co/datasets/aimosprite/moh_8_rollouts_tool_calls.
