arjhinety/OpenGrad-ToolPolicy-Canonical-v2
This is a provenance-preserving canonical candidate corpus. It is a pre-training canonical release, not an empirically selected or recommended training mixture. What this release is OpenGrad ToolPolicy Canonical v2 is a provenance-preserving, model-independent normalization of public tool-use datasets in which every record declares what it supervises. It exists because not every legitimate post-training corpus has the same conversational trajectory shape, and discarding a… See the full description on the dataset page: https://huggingface.co/datasets/arjhinety/OpenGrad-ToolPolicy-Canonical-v2.
fix: the licence file still described the three-source partial snapshot
fix: the licence file still described the three-source partial snapshot
Correct card against committed artifacts (OpenGrad claim audit, 2026-09-13)
OpenGrad ToolPolicy Canonical v2 (final): 4 sources, 173,237 records, explicit supervision contracts on every record
initial commit
