ssaraf1/slm-workflow-planner-policy-v2
SLM Workflow Planner — Policy-Corrected Instruction Tuning Dataset (v2) Overview High-quality instruction-tuning dataset for training a Small Language Model (SLM) to serve as a workflow execution planner. The model learns to make policy-aware decisions about workflow transitions: when to proceed (NEXT), retry (RETRY), parallelize (FORK), synchronize (JOIN), or escalate (META). Key Features 648K instruction pairs across 2 stages (decision type +… See the full description on the dataset page: https://huggingface.co/datasets/ssaraf1/slm-workflow-planner-policy-v2.
This repository belongs to ssaraf1 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
