CoolFace
Datasetpublic

MauroPello/reasoning-gym-verl-datasets

reasoning-gym-verl-datasets This dataset contains procedurally generated reasoning tasks from the Reasoning Gym (r-gym) framework, structured and pre-processed in parquet format for training models with veRL. These datasets were used to train MauroPello/Qwen3-1.7B-RL-final using GRPO (Group Relative Policy Optimization). Dataset Splits & Structure Split Name Path Size (Examples) Description train train.parquet 100,000 Raw training set containing… See the full description on the dataset page: https://huggingface.co/datasets/MauroPello/reasoning-gym-verl-datasets.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
0likes85downloads
settings

This repository belongs to MauroPello on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namereasoning-gym-verl-datasets
visibilitypublic
licenceapache-2.0
gatedno
ownerMauroPello
Account settings
MauroPello/reasoning-gym-verl-datasets · CoolFace