easyr1
EasyR1-qwen25vl-7b-spar234k-59609323-step850-mergedqwen2_5vl_7b_easyr1_114k_large_runEasyR1-qwen25vl-7b-spar234k-sgrpo-prob0-half0-step20qwen2_5vl_7b_easyr1_114k_reproduceqwen2_5vl_7b_easyr1_10k_hard_qwen7b_easy_gta1_4MP_no_resolution_in_prompt_lr_1_0e-06_bs16_z3qwen2_5vl_7b_easyr1_114k_reproduce_z2EasyR1-qwen25vl-7b-vgllm-spar234k-rl-multi-depth-17126318-step10qwen2_5vl_7b_easyr1_20k_hard_qwen7b_easy_gta1_4MP
Datasets
All datasets matching “easyr1”easyr1-grounding-dataset-30k-not_grounded-SE-GUI-3B-2MPEasyR1-qwen3vl-rl
EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework
Used by Amazon Web Services
This project is a clean fork of the original veRL project to support vision language models, we thank all the authors for providing such a high-performance RL training framework.
EasyR1 is efficient and scalable due to the design of HybirdEngine and the latest release of vLLM's SPMD mode.
Features
Supported models
Llama3/Qwen2/Qwen2.5/Qwen3 language… See the full description on the dataset page: https://huggingface.co/datasets/johnschaefer/EasyR1-qwen3vl-rl.easyr1-103k-4MP-jedi-ui-vision-gta1-data
easyr1-103k-4MP-jedi-ui-vision-gta1-data
Merged dataset composed of the following sources:
datasets/easyr1-63k-nores-jedi-fix-synced-ui-vision-manually-labeled-icon-data-from-yt-4MP (63031 samples in split train)
datasets/easyr1-grounding-gta1-4MP-easy-qwen7b-hard-gta1-7b (39943 samples in split train)
Summary
Generated on: 2025-09-18 06:29:16 UTC
Split: train
Column strategy: intersection
Samples after merge: 102974
Usage
from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-cua-dev/easyr1-103k-4MP-jedi-ui-vision-gta1-data.qwen3-resize-easyr1-110k-bbox0p05-remove-pixmo-uground-seeclickeasyr1-114k-hard-qwen7b-easy-gta1-4MP-nores-jedi-fix-synced-aug-jitter-tokenizedeasyr1-103k-4MP-jedi-ui-vision-gta1-data-bbox-filtered-0p05
