CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nvidia /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1 Dataset Description: We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838 different… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.tabular10K<n<100K32 likes1.4k downloads7mo agoHugging Face02llamafactory /reason-tool-use-demo-1500 Dataset info The dataset is a selection of reasoning toolcalls data from https://huggingface.co/datasets/interstellarninja/hermes_reasoning_tool_use, which contains data from Hermes-Tools、Glaive-FC、ToolAce、Nvidia-When2Call. The format has been transformed to adapt llama-factory v1 training pipeline. textquestion-answering1K<n<10K1 likes1.1k downloads9mo agoHugging Face03rmems /browser-tool-use-trajectories Browser Tool Use Trajectories Rights & intended use: legacy public research corpus / portfolio artifact. Hosted frontier-model outputs are research-only inputs under project policy (synthetic-factory#161): intended_use: research_only, project_training_policy: blocked. Not training data for any model-weight update. Machine-readable record: rights.json. Release status: The raw, uncurated payload is now published under data/raw/. It is available for inspection and… See the full description on the dataset page: https://huggingface.co/datasets/rmems/browser-tool-use-trajectories.text1K<n<10K1 likes770 downloads22d agoHugging Face04spade-rl /SPADE-Environment-Pool-GPT5.5-ToolUse SPARE GPT-5.5 Multi-Turn Tool-Use Games v1 A public static pool of 11,039 validated multi-turn tool-use environments generated by GPT-5.5 for SPARE actor training. Training alignment Source recipe: Qwen3-30B-A3B 0624 tool-use GAMES configuration 400 rollouts x 24 games/rollout = 9,600 no-reuse games required 11,039 validated games provide 1,439 games of headroom Six balanced skills: API orchestration, data retrieval, state modification, error recovery, tool… See the full description on the dataset page: https://huggingface.co/datasets/spade-rl/SPADE-Environment-Pool-GPT5.5-ToolUse.tabularreinforcement-learning10K<n<100K1 likes609 downloads28d agoHugging Face05msr-spare-1 /qwen3-4b-0701-tooluse-glory-kl0-spare-games-envs qwen3-4B-Instruct-0701-tooluse-glory-kl0 — generated environments Environments generated by the SPARE proposer during training run 223t1pws (qwen3-4B-Instruct-0701-tooluse-glory-kl0), recovered from the spare-viz durable cache. The run's scratch directory no longer exists; this dataset is the surviving copy. Games 350 Steps covered 16 (step 0–384) With recovered skill 350 With hint 0 Actor / proposer model… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-4b-0701-tooluse-glory-kl0-spare-games-envs.textn<1K0 likes206 downloads1mo agoHugging Face06evoeval /EvoEval_tool_usetextn<1K4 likes187 downloads2y agoHugging Face07spade-rl /SPADE-Grounding-Corpus-ToolUse-15K SPADE grounding corpus: tool use (15k) Reference documents the SPADE Environment Designer is grounded on when generating multi-turn tool-use environments. 15,552 source files drawn from nvidia/Nemotron-Pretraining-Code-v3. Documents 15,552 Setting tool_use Fields text (the document), metadata (source provenance) Each generation prompt embeds one sampled document, so the environments a Designer writes stay anchored to a real concept or technique rather than… See the full description on the dataset page: https://huggingface.co/datasets/spade-rl/SPADE-Grounding-Corpus-ToolUse-15K.texttext-generation10K<n<100K1 likes139 downloads28d agoHugging Face08rmems /tool-use-preference-pairs Tool Use Preference Pairs Rights & intended use: legacy public research corpus / portfolio artifact. Hosted frontier-model outputs are research-only inputs under project policy (synthetic-factory#161): intended_use: research_only, project_training_policy: blocked. Not training data for any model-weight update. Machine-readable record: rights.json. Release status: The raw, uncurated payload is now published under data/raw/. It is available for inspection and reproducibility… See the full description on the dataset page: https://huggingface.co/datasets/rmems/tool-use-preference-pairs.text1K<n<10K0 likes130 downloads22d agoHugging Face09protogonos /verified-tool-use-dataset Verified tool-use trajectories for LLM agents This was a time-boxed experiment by an autonomous agent (Protogonos), now concluded. Nothing here is offered for sale or for hire, and no payment is accepted. Multi-turn function-calling conversations for training and evaluating tool-using agents — 48 trajectories across 16 domains, with every tool call checked against its tool's JSON-Schema. The free sample in this repo is a real slice of the full set: the viewer above renders it… See the full description on the dataset page: https://huggingface.co/datasets/protogonos/verified-tool-use-dataset.texttext-generationn<1K1 likes117 downloads24d agoHugging Face10schneiderkamplab /dfm11-toolace-native-tool-use-repaired dfm11-toolace-native-tool-use-repaired ToolACE conversations with declared-name parsing and complete parallel result binding. This is a DFM11 replacement for schneiderkamplab/dfm10-toolace-native-tool-use. All rows pass exhaustive structural validation. See metadata/manifest.json. text10K<n<100K0 likes99 downloads19d agoHugging Face11msr-spare-1 /SPADE-Environments-Qwen3-30B-ToolUse qwen3-30B-A3B-Instruct-0703-tooluse-glory-kl005 — generated environments Environments generated by the SPARE proposer during training run 2hjdrbeh (qwen3-30B-A3B-Instruct-0703-tooluse-glory-kl005), recovered from the spare-viz durable cache. The run's scratch directory no longer exists; this dataset is the surviving copy. Games 260 Steps covered 7 (step 0–192) With recovered skill 260 With hint 0 Actor / proposer model… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/SPADE-Environments-Qwen3-30B-ToolUse.textn<1K0 likes77 downloads1mo agoHugging Face12asingh15 /qwen35-2b-tool-use-qwen36-27b-curation-candidates Full candidate collections: 2B tool use + 27B data curation This public Dataset contains two complete, unredacted, exact-40 candidate collections: Tool use: Qwen/Qwen3.5-2B at 15852e8c16360a2fea060d615a32b45270f8a8fc, 5,849 tasks and 233,960 candidates across ACEBench, APIBank, BFCL, BIRD, NESTFUL, Spider, and TravelPlanner. Data curation: Qwen/Qwen3.6-27B at 6a9e13bd6fc8f0983b9b99948120bc37f49c13e9, 5,021 targets and 200,840 candidates, plus the source target rows and the… See the full description on the dataset page: https://huggingface.co/datasets/asingh15/qwen35-2b-tool-use-qwen36-27b-curation-candidates.tabulartext-generation100K<n<1M0 likes77 downloads1mo agoHugging Face13BertilBraun /voice-light-tool-use-synthetic Voice Light Teacher-Led Tool-Use Synthetic This repository contains the current canonical synthetic source dataset for Voice Light's conversational tool-use fine-tuning. The current revision contains 3,994 provider-neutral English conversations generated from 4,000 deterministic teacher-led scenario plans. Every conversation has four user turns so follow-up requests can depend naturally on prior turns and tool results. The Hugging Face train split names the canonical JSONL file… See the full description on the dataset page: https://huggingface.co/datasets/BertilBraun/voice-light-tool-use-synthetic.texttext-generation1K<n<10K0 likes74 downloads2mo agoHugging Face14elikoy /deepseek-v41-flash-thinking-toolusetextn<1K0 likes67 downloads7d agoHugging Face15guanaco /Tool_usetextn<1K3 likes51 downloads3y agoHugging Face16ajibawa-2023 /Test-Tool-Use-AgentThis is in Test mode. Don't use for the time being. Seed Data : glaiveai/reasoning-v1-20m LLM Used : Qwen3.6-27B text10K<n<100K1 likes50 downloads3mo agoHugging Face17stindardlogic /agentic-dpo-tool-use-3k Agentic DPO Tool-Use Pairs (3K) Synthetic DPO preference pairs for training LLMs to use tools correctly in agentic settings. Dataset Description 3,000 preference pairs covering 7 tool categories: web_search — real-time web search calculator — mathematical expression evaluation weather_api — current weather retrieval code_interpreter — Python code execution database_query — SQL database queries stock_price — financial data lookup translate — multilingual… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/agentic-dpo-tool-use-3k.texttext-generation1K<n<10K0 likes50 downloads2mo agoHugging Face18Makeen-AraFC /hermes-tool-use-reasoning-ar Arabic Hermes Tool-Use Reasoning Arabic translation of the Hermes Tool Use Reasoning dataset for research on Arabic function calling, tool selection, argument generation, and tool-call verification. The release contains 2,422 examples covering 1,172 unique tools in ShareGPT format. Dataset Structure Each example contains: { "tools": [...], "conversations": [...] } tools: candidate tool declarations, including names, descriptions, parameter names, types, and… See the full description on the dataset page: https://huggingface.co/datasets/Makeen-AraFC/hermes-tool-use-reasoning-ar.texttext-generation1K<n<10K0 likes38 downloads1mo agoHugging Face19stindardlogic /tool-use-dpo-100k Tool Use DPO (100K) 100,000 DPO preference pairs for training models to make correct tool use decisions. Each pair includes a user prompt, a chosen response that correctly reasons about tool use, and a rejected response that makes a tool use mistake. Covers 6 decision categories and 23 distinct tool use failure patterns found in production agentic AI systems. Motivation As LLMs are deployed in agentic pipelines with access to tools (APIs, databases, code execution… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/tool-use-dpo-100k.texttext-generation100K<n<1M0 likes37 downloads2mo agoHugging Face20Makeen-AraFC /glaive-tool-use-reasoning-ar Arabic Glaive Tool-Use Reasoning Arabic translation and augmentation of the Glaive Function Calling data for research on Arabic function calling, tool selection, argument generation, and tool-call verification. The release contains 3,336 examples covering 414 tools in ShareGPT format. Dataset Structure Each example contains: { "tools": [...], "conversations": [...] } tools: candidate tool declarations, including names, descriptions, parameter names, types… See the full description on the dataset page: https://huggingface.co/datasets/Makeen-AraFC/glaive-tool-use-reasoning-ar.texttext-generation1K<n<10K0 likes37 downloads1mo agoHugging Face21asingh15 /qwen35-2b-tool-use-candidates Qwen3.5-2B Full Tool-Use Candidates This is the complete certified seven-suite tool-use collection for Qwen/Qwen3.5-2B at immutable model revision 15852e8c16360a2fea060d615a32b45270f8a8fc. 5,849 original tasks exactly 40 unprivileged candidates per task 233,960 complete candidate responses ACEBench, APIBank, BFCL, BIRD, NESTFUL, Spider, and TravelPlanner AppWorld is not included data/unprivileged.jsonl is a byte-for-byte copy of the certified collection. Original task IDs… See the full description on the dataset page: https://huggingface.co/datasets/asingh15/qwen35-2b-tool-use-candidates.texttext-generation1K<n<10K0 likes34 downloads1mo agoHugging Face22Vendex /agentic-tooluse-computer-browsertextn<1K0 likes33 downloads2mo agoHugging Face23asingh15 /qwen36-27b-tool-use-candidates Qwen3.6-27B Full Tool-Use Candidates This is the complete certified seven-suite tool-use collection for Qwen/Qwen3.6-27B at revision 6a9e13bd6fc8f0983b9b99948120bc37f49c13e9. 5,849 original tasks exactly 40 unprivileged candidates per task 233,960 complete candidate responses ACEBench, APIBank, BFCL, BIRD, NESTFUL, Spider, and TravelPlanner AppWorld is not included data/unprivileged.jsonl is a byte-for-byte copy of the certified collection. Original task IDs, task text, tool… See the full description on the dataset page: https://huggingface.co/datasets/asingh15/qwen36-27b-tool-use-candidates.texttext-generation1K<n<10K0 likes32 downloads1mo agoHugging Face24Arsh9210 /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1 Dataset Description: We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838… See the full description on the dataset page: https://huggingface.co/datasets/Arsh9210/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.tabular10K<n<100K0 likes31 downloads2mo agoHugging Face25Mgmgrand420 /Agent-Tool-Use-Dialogue-Open-Dataset Open Agent RL Dataset: High Quality AI Agent | Tool Use & Function Calls | Reinforcement Learning Datasets Github|Huggingface|Pypi | Open Source AI Agent Marketplace DeepNLP|Agent RL Dataset DeepNLP website provides high quality, genuine, online users' request of Agent & RL datasets to help LLM foundation/SFT/Post Train to get more capable models at function call, tool use and planning. The datasets are collected and sampled from users' requests on our various clients (Web/App/Mini… See the full description on the dataset page: https://huggingface.co/datasets/Mgmgrand420/Agent-Tool-Use-Dialogue-Open-Dataset.textn<1K0 likes30 downloads8mo agoHugging Face26monodox /agent-conversations-and-tool-usetextn<1K0 likes24 downloads5mo agoHugging Face27goodknightleo /mythos-tooluse-hermes-suitetext10K<n<100K1 likes23 downloads3mo agoHugging Face28msr-spare-1 /qwen3-4b-0630-tooluse-eval-aligned-r32-spare-games-envs qwen3-4B-Instruct-0630-tooluse-eval-aligned-r32 — generated environments Environments generated by the SPARE proposer during training run 050mlekj (qwen3-4B-Instruct-0630-tooluse-eval-aligned-r32), recovered from the spare-viz durable cache. The run's scratch directory no longer exists; this dataset is the surviving copy. Games 456 Steps covered 21 (step 0–448) With recovered skill 456 With hint 0 Actor / proposer model… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-4b-0630-tooluse-eval-aligned-r32-spare-games-envs.textn<1K0 likes20 downloads1mo agoHugging Face29LorthGyu /indonesian-agent-tooluse Agent Tool-Use Bahasa Indonesia 🤖 Dataset 241 contoh function-calling / tool-use berbahasa Indonesia — instruksi user natural + tool definitions + tool calls yang tepat + response. Kenapa dataset ini ada? Tool-calling adalah tren paling panas di HF (orca-agentinstruct 466 likes, Toucan 226, DeepScaleR 205) — tapi tidak ada satu pun dataset tool-use berbahasa Indonesia. Model lokal yang bisa panggil tool (cek cuaca, booking, cari rute) dalam bahasa Indonesia = gap… See the full description on the dataset page: https://huggingface.co/datasets/LorthGyu/indonesian-agent-tooluse.texttext-generationn<1K0 likes18 downloads2mo agoHugging Face30msr-spare-1 /qwen3-4b-0701-tooluse-glory-kl005-spare-games-envs qwen3-4B-Instruct-0701-tooluse-glory-kl005 — generated environments Environments generated by the SPARE proposer during training run ny5pjmw9 (qwen3-4B-Instruct-0701-tooluse-glory-kl005), recovered from the spare-viz durable cache. The run's scratch directory no longer exists; this dataset is the surviving copy. Games 450 Steps covered 22 (step 0–449) With recovered skill 450 With hint 0 Actor / proposer model… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-4b-0701-tooluse-glory-kl005-spare-games-envs.textn<1K0 likes18 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.