cy-330/tau2-telecom-agent-sft
τ²-bench telecom — teacher trajectories for agent SFT 784 accepted multi-turn tool-use trajectories on the telecom domain of τ²-bench, collected to cold-start an 8B model before reinforcement learning. Training code, the full lab record and the RL stages that follow are at yuecao365/tau2telecom_RL. The point of this domain is dual control: the agent has thirteen backend APIs, the customer has thirty tools on their own handset, and 76% of the actions a task expects can only be… See the full description on the dataset page: https://huggingface.co/datasets/cy-330/tau2-telecom-agent-sft.
063
Correct the per-task count and link back to the training repository
Add 784 teacher trajectories for tau2-bench telecom agent SFT
initial commit
