CoolFace
Datasetpublic

AmanPriyanshu/tool-reasoning-sft-TOOLS-toucan-1.5m-sft-tool-use-data-cleaned-rectified-333k

Toucan - OSS High Quality (Hermes Reasoning Format) Filtered and restructured subset of Agent-Ark/Toucan-1.5M. Format Inspiration: SupritiVijay/dr-tulu-sft-deep-research-agent-data-cleaned-rectified Filters applied: OSS split only · overall_score > 3.0 · valid role transitions only Size: ~333K examples Format Each example is a multi-turn conversation with strict role transitions: system → user → reasoning → tool_call → tool_output → reasoning → ... → answer… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-TOOLS-toucan-1.5m-sft-tool-use-data-cleaned-rectified-333k.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
0likes64downloads
Dataset Card

Toucan - OSS High Quality (Hermes Reasoning Format)

Filtered and restructured subset of Agent-Ark/Toucan-1.5M. Format Inspiration: SupritiVijay/dr-tulu-sft-deep-research-agent-data-cleaned-rectified

Filters applied: OSS split only · overall_score > 3.0 · valid role transitions only

Size: ~333K examples


Format

Each example is a multi-turn conversation with strict role transitions:

system → user → reasoning → tool_call → tool_output → reasoning → ... → answer
RoleContent
systemTool schemas + instructions
userQuestion
reasoning<think>...</think>
tool_call<tool_call>{"name": ..., "arguments": {...}}</tool_call>
tool_output<tool_response>...</tool_response>
answer<answer>...</answer>

Multi-turn conversations follow answer → user transitions.


Changes from Original

  • —Mapped reasoning_content → <think> blocks
  • —Parsed tool call JSON; stripped call_id and null schema fields
  • —Inserted synthetic <think> bridges where tool_output → answer transitions were missing reasoning
  • —Wrapped final answers in <answer> tags
  • —Dropped rows with invalid transitions (~1.2%)

Original dataset: Agent-Ark/Toucan-1.5M - Apache 2.0