datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Nemotron-Cascade-2-RL-data
Dataset Description:
The Nemotron-Cascade-2-RL dataset is a curated reinforcement learning (RL) dataset blend used to train Nemotron-Cascade-2-30B-A3B model. It includes instruction-following RL, multi-domain RL, on-policy distillation, and software engineering RL (SWE-RL) data.
This dataset is ready for commercial use.
The dataset contains the following subset:
IF-RL
Contains 45,879 training samples for instruction-following RL. Our curation process mainly… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-Cascade-2-RL-data.Nemotron-Cascade-RL-RLHF
Dataset Description:
The Nemotron-Cascade-RL-RLHF dataset is designed for Reinforcement Learning from Human Feedback (RLHF) training. It contains prompts and associated metadata to support the development of language model alignment.
This dataset is ready for commercial use.
The dataset contains the following subset:
RLHF Training Data
This data contains 45,882 samples used for RLHF training. It includes prompts, data sources, and category information.
This dataset is a… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-Cascade-RL-RLHF.rlve-rollouts-nemotron-cascade-8b-qwen3-1.7bNemotron-Cascade-RL-RLHF
