CoolFace
Datasetpublic

ssaraf1/slm-workflow-planner-policy-v2

SLM Workflow Planner — Policy-Corrected Instruction Tuning Dataset (v2) Overview High-quality instruction-tuning dataset for training a Small Language Model (SLM) to serve as a workflow execution planner. The model learns to make policy-aware decisions about workflow transitions: when to proceed (NEXT), retry (RETRY), parallelize (FORK), synchronize (JOIN), or escalate (META). Key Features 648K instruction pairs across 2 stages (decision type +… See the full description on the dataset page: https://huggingface.co/datasets/ssaraf1/slm-workflow-planner-policy-v2.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes35downloads

ssaraf1/slm-workflow-planner-policy-v2 · main · files are served by the source, never re-hosted here