datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
embodied_reasoner
Embodied-Reasoner Dataset
Dataset Overview
Embodied-Reasoner is a multimodal reasoning dataset designed for embodied interactive tasks. It contains 9,390 Observation-Thought-Action trajectories for training and evaluating multimodal models capable of performing complex embodied tasks in indoor environments.
Key Features
📸 Rich Visual Data: Contains 64,000 first-person perspective interaction images🤔 Deep Reasoning Capabilities: 8 million thought… See the full description on the dataset page: https://huggingface.co/datasets/zwq2018/embodied_reasoner.orz_math_57k_collection
Open Reasoner Zero
An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Paper Arxiv Link 👁️
Overview 🌊
We introduce Open-Reasoner-Zero, the first open source implementation of large-scale reasoning-oriented RL training focusing on scalability, simplicity and accessibility.
To enable broader participation in this pivotal moment we witnessed and accelerate research towards artificial general intelligence (AGI)… See the full description on the dataset page: https://huggingface.co/datasets/Open-Reasoner-Zero/orz_math_57k_collection.VerMultiThis repository contains the data presented in LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.
Project page: https://forjadeforest.github.io/LMM-R1-ProjectPage
verl_format_batched_splits_v1orz_math_13k_collection_hard
Open Reasoner Zero
An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Paper Arxiv Link 👁️
Overview 🌊
We introduce Open-Reasoner-Zero, the first open source implementation of large-scale reasoning-oriented RL training focusing on scalability, simplicity and accessibility.
To enable broader participation in this pivotal moment we witnessed and accelerate research towards artificial general intelligence (AGI)… See the full description on the dataset page: https://huggingface.co/datasets/Open-Reasoner-Zero/orz_math_13k_collection_hard.ui_design_reasoner_nemotronvlDynaMath_Evalui_design_reasoner_rollouts_nemotronvlmathverse_1k_valverl_format_test_v2Data_ver_0.1
