CoolFace
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nvidia /Nemotron-RL-Agentic-Function-Calling-Pivot-v1 Dataset Description: This is a RL dataset for general function-calling by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection of… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Function-Calling-Pivot-v1.text1K<n<10K14 likes2.1k downloads7mo agoHugging Face02nvidia /Nemotron-RL-Agentic-Terminal-Pivot-v1 Dataset Description The Nemotron-RL-Agentic-Terminal-Pivot-v1 dataset provides training samples for reinforcement learning of command-line ("terminal use") LLM agents with the terminus_judge environment in NeMo Gym. Each record is a single agent decision point extracted from a successful agent trajectory on a terminal task: responses_create_params.input — the prompt: the task instruction plus the terminal interaction history (prior agent actions and terminal outputs) up to the… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Terminal-Pivot-v1.texttext-generation10K<n<100K31 likes1.7k downloads27d agoHugging Face03nvidia /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1 Dataset Description: We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838 different… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.tabular10K<n<100K32 likes1.4k downloads7mo agoHugging Face04nvidia /Nemotron-RL-Agentic-SWE-Pivot-v1 Dataset Description: The SWE-RL dataset provides GitHub issues for training and validating real-world software engineering agents using the OpenHands environment in NeMo Gym. The dataset is a refactored version of the SWE-Gym and R2E-Gym datasets to support the NeMo Gym input format. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection of training… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-SWE-Pivot-v1.tabular10K<n<100K15 likes1.2k downloads3mo agoHugging Face05open-athena /glm52-datagen-r11-100-agentic-function-calling-pivot-v2-tracestext1K<n<10K1 likes288 downloads1mo agoHugging Face06wckwan /WebShop-Qwen3-8B-Adaptive-Pivot-Fenced-s1000-evaltextn<1K0 likes106 downloads4h agoHugging Face07wckwan /WebShop-Qwen3-8B-Adaptive-Pivot-analysistext1K<n<10K0 likes85 downloads13d agoHugging Face08jamesdborin /Nemotron-RL-Agentic-SWE-Pivot-v1-prompt-only Nemotron-RL-Agentic-SWE-Pivot-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-SWE-Pivot-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts. null_or_empty_rows.md: row indexes where… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-SWE-Pivot-v1-prompt-only.tabular10K<n<100K0 likes77 downloads3mo agoHugging Face09jamesdborin /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1-prompt-only Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1-prompt-only.tabular10K<n<100K0 likes72 downloads3mo agoHugging Face10Dabou /Nemotron-RL-Agentic-Terminal-Pivot-v1 Dataset Description The Nemotron-RL-Agentic-Terminal-Pivot-v1 dataset provides training samples for reinforcement learning of command-line ("terminal use") LLM agents with the terminus_judge environment in NeMo Gym. Each record is a single agent decision point extracted from a successful agent trajectory on a terminal task: responses_create_params.input — the prompt: the task instruction plus the terminal interaction history (prior agent actions and terminal outputs) up to the… See the full description on the dataset page: https://huggingface.co/datasets/Dabou/Nemotron-RL-Agentic-Terminal-Pivot-v1.texttext-generation10K<n<100K0 likes65 downloads27d agoHugging Face11jamesdborin /Nemotron-RL-Agentic-Function-Calling-Pivot-v1-prompt-only Nemotron-RL-Agentic-Function-Calling-Pivot-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-Function-Calling-Pivot-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-Function-Calling-Pivot-v1-prompt-only.tabular1K<n<10K0 likes53 downloads3mo agoHugging Face12Bobollinix /Nemotron-RL-Agentic-SWE-Pivot-v1 Dataset Description: The SWE-RL dataset provides GitHub issues for training and validating real-world software engineering agents using the OpenHands environment in NeMo Gym. The dataset is a refactored version of the SWE-Gym and R2E-Gym datasets to support the NeMo Gym input format. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection of training… See the full description on the dataset page: https://huggingface.co/datasets/Bobollinix/Nemotron-RL-Agentic-SWE-Pivot-v1.tabular10K<n<100K0 likes31 downloads5d agoHugging Face13Arsh9210 /Nemotron-RL-Agentic-SWE-Pivot-v1 Dataset Description: The SWE-RL dataset provides GitHub issues for training and validating real-world software engineering agents using the OpenHands environment in NeMo Gym. The dataset is a refactored version of the SWE-Gym and R2E-Gym datasets to support the NeMo Gym input format. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection of training… See the full description on the dataset page: https://huggingface.co/datasets/Arsh9210/Nemotron-RL-Agentic-SWE-Pivot-v1.tabular10K<n<100K0 likes28 downloads2mo agoHugging Face14Arsh9210 /Nemotron-RL-Agentic-Function-Calling-Pivot-v1 Dataset Description: This is a RL dataset for general function-calling by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection… See the full description on the dataset page: https://huggingface.co/datasets/Arsh9210/Nemotron-RL-Agentic-Function-Calling-Pivot-v1.text1K<n<10K0 likes27 downloads2mo agoHugging Face15Arsh9210 /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1 Dataset Description: We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838… See the full description on the dataset page: https://huggingface.co/datasets/Arsh9210/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.tabular10K<n<100K0 likes23 downloads2mo agoHugging Face16alucent /mirror-Nemotron-RL-Agentic-Function-Calling-Pivot-v1gated Dataset Description: This is a RL dataset for general function-calling by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection… See the full description on the dataset page: https://huggingface.co/datasets/alucent/mirror-Nemotron-RL-Agentic-Function-Calling-Pivot-v1.text1K<n<10K0 likes17 downloads2mo agoHugging Face17open-athena /nemotron-gym-agentic-swe-pivot laion/nemotron-gym-agentic-swe-pivot Harbor task-binary dataset (3,978 tasks) converted from nvidia/Nemotron-RL-Agentic-SWE-Pivot-v1 (part of the nvidia/Nemotron-Post-Training-v3 collection). Each row is a valid Harbor task binary: columns path (str) and task_binary (gzip tar). Converted with the OpenThoughts-Agent data.nemotron_gym framework. Grading: Single-step SWE tool-call match (case-sensitive, whitespace-normalized). texttext-generation1K<n<10K0 likes13 downloads20d agoHugging Face18IvanMiao /ancientChinese-pivot-FR_MTtextn<1K0 likes10 downloads9mo agoHugging Face19tcapelle /pivoted-trustworthy-alignmenttext1K<n<10K0 likes7 downloads2y agoHugging Face20CodeClarity /PivotTranslationstextn<1K0 likes7 downloads7mo agoHugging Face21Gopher-Lab /bankless_ROLLUP_Fed_Pivot__Ledger_Scare__New_Warren_Bill__Airdrop_Szntextn<1K0 likes6 downloads2y agoHugging Face22Mayur295 /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1 Dataset Description: We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838… See the full description on the dataset page: https://huggingface.co/datasets/Mayur295/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.tabular10K<n<100K0 likes6 downloads3mo agoHugging Face23alucent /mirror-Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1gated Dataset Description: We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838… See the full description on the dataset page: https://huggingface.co/datasets/alucent/mirror-Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.tabular10K<n<100K0 likes6 downloads2mo agoHugging Face24open-athena /nemotron-gym-agentic-conversational-tool-use-pivot laion/nemotron-gym-agentic-conversational-tool-use-pivot Harbor task-binary dataset (96,965 tasks) converted from nvidia/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1 (part of the nvidia/Nemotron-Post-Training-v3 collection). Each row is a valid Harbor task binary: columns path (str) and task_binary (gzip tar). Converted with the OpenThoughts-Agent data.nemotron_gym framework. Grading: Single-step: tool-call match (function_call) / LLM judge (message). texttext-generation10K<n<100K0 likes6 downloads20d agoHugging Face25open-athena /nemotron-gym-agentic-function-calling-pivot laion/nemotron-gym-agentic-function-calling-pivot Harbor task-binary dataset (9,579 tasks) converted from nvidia/Nemotron-RL-Agentic-Function-Calling-Pivot-v1 (part of the nvidia/Nemotron-Post-Training-v3 collection). Each row is a valid Harbor task binary: columns path (str) and task_binary (gzip tar). Converted with the OpenThoughts-Agent data.nemotron_gym framework. Grading: Single-step: tool-call match (function_call) / LLM judge (message). texttext-generation1K<n<10K0 likes6 downloads20d agoHugging Face26pskulkarni /lad-extended-pivotinggated LAD Extended Pivoting Dataset Synthetic multi-turn conversations with deliberately extended pivoting phases (4+ pivoting turns) for validating early adversarial detection. Companion to lad-multiturn-adversarial. Why This Dataset? The core LAD dataset has a mean of 3.3 pivoting turns per adversarial conversation, yielding 22-26% early detection. This dataset tests whether longer pivoting phases improve early detection — they do dramatically: Metric Core… See the full description on the dataset page: https://huggingface.co/datasets/pskulkarni/lad-extended-pivoting.texttext-classificationn<1K0 likes4 downloads6mo agoHugging Face27syazayacob /crop_data_pivot_logtabular100K<n<1M0 likes3 downloads1y agoHugging Face28syazayacob /crop_data_pivottabular100K<n<1M0 likes2 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.