verl
FC-VERL-JSON-1.5B-GGUFptdbench-verl-implementation-torch-functionalsemsimula-fock-parflm-anisogaussian-vtheta-owt-d384-verlet-instabilityptdbench-verl-coding-task-evaluatorptdbench-verl-coding-tasks-function-callverl-grpo-medium-qwen3-4b-step129-reproverl-grpo-medium-qwen3-4b-step129verl-grpo-medium-qwen3-4b-step100
Datasets
All datasets matching “verl”waterbird-verlverl_vla_libero_collectedThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "panda",
"total_episodes": 32,
"total_frames": 2990,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 1e-06,
"fps": 10,
"splits": {
"train": "0:32"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Miical/verl_vla_libero_collected.verl_mask_training
👋 Hi, everyone!
verl is a RL training library initiated by ByteDance Seed team and maintained by the verl community.
verl: Volcano Engine Reinforcement Learning for LLMs
verl is a flexible, efficient and production-ready RL training library for large language models (LLMs).
verl is the open-source version of HybridFlow: A Flexible and Efficient RLHF Framework paper.
verl is flexible and easy to use with:
Easy extension of diverse RL algorithms: The… See the full description on the dataset page: https://huggingface.co/datasets/SeanWang0027/verl_mask_training.deepscaler-preview-verl
DeepScaleR-Preview VERL
📊 Dataset Summary
This dataset contains 35,789 mathematical reasoning problems in VERL format, processed from agentica-org/DeepScaleR-Preview-Dataset.
Key Features:
35,789 high-quality math problems
Converted to VERL format for reward modeling
Verified ground truth answers
Ready for reinforcement learning training
🔗 Source Dataset
Original Repository
Repository:… See the full description on the dataset page: https://huggingface.co/datasets/sungyub/deepscaler-preview-verl.DeepScaleR-Preview-Dataset-verl-formatgsm8k-v0.4.1The dataset is generated based on verl 0.4.1 with command:
python3 examples/data_preprocess/gsm8k.py
