datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
rlvr-code-data-python-r1-format-filteredMagpie-Reasoning-V2-250K-CoT-Deepseek-R1-Llama-70B-formattedaime_1983_2023_grok-3-mini-high_traces_r1_formattedwildjailbreak-r1-v2-format-filteredMMMU-LLM-R1-format
MMMU-LLM-R1 Reformatted Dataset
Synthetic-Japanese-Roleplay-SFW-DeepSeek-R1-0528-10k-formatted
Synthetic-Japanese-Roleplay-SFW-DeepSeek-R1-0528-10k-formatted
概要
deepseek-ai/DeepSeek-R1-0528を用いて作成した日本語ロールプレイデータセットであるAratako/Synthetic-Japanese-Roleplay-SFW-DeepSeek-R1-0528-10kにsystem messageを追加して整形したデータセットです。
データの詳細については元データセットのREADMEを参照してください。
ライセンス
MITライセンスの元配布します。
rlvr-code-data-python-r1-format-filteredwildchat-r1-p2-format-filteredllava-cot-100k-r1-format
llava-cot-100k-r1-format: A dataset for Vision Reasoning GRPO Training
Images
Images data can be access from https://huggingface.co/datasets/Xkev/LLaVA-CoT-100k
SFT dataset
https://huggingface.co/datasets/di-zhang-fdu/R1-Vision-Reasoning-Instructions
Citations
@misc {di_zhang_2025,
author = { {Di Zhang} },
title = { llava-cot-100k-r1-format (Revision 87d607e) },
year = 2025,
url = {… See the full description on the dataset page: https://huggingface.co/datasets/di-zhang-fdu/llava-cot-100k-r1-format.the-algorithm-python-r1-format-filteredMM-MathInstruct-to-r1-format-filtered
MM-MathInstruct-to-r1-format-filtered
MM-MathInstruct dataset transformed to R1 format and filtered by token length and image quality
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized input… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/MM-MathInstruct-to-r1-format-filtered.wildchat-r1-p2-filtered-formatpersona-precise-if-r1-merged-format-filtereddefault-open-r1-math-90k-formatThis dataset, derived from the default version of OpenR1-Math-220K, has been reformatted and organized for improved usability and model training. The following data processing and quality filtering measures were implemented:
Data Processing and Quality Filtering Methodology
Structural Integrity Validation:
Ensures data consistency by verifying equal lengths across correctness_math_verify, is_reasoning_complete, and generations lists.
Confirms the presence of all required fields within each… See the full description on the dataset page: https://huggingface.co/datasets/DylanDDeng/default-open-r1-math-90k-format.aime_1983_2023_grok-3-mini-high_traces_16392_r1_formattedaya-100k-r1-format-filteredtulu_v3.9_wildchat_100k_english-r1-format-filteredSynthetic-Japanese-Roleplay-NSFW-DeepSeek-R1-0528-10k-formatted
Synthetic-Japanese-Roleplay-NSFW-DeepSeek-R1-0528-10k-formatted
概要
deepseek-ai/DeepSeek-R1-0528を用いて作成した日本語ロールプレイデータセットであるAratako/Synthetic-Japanese-Roleplay-NSFW-DeepSeek-R1-0528-10kにsystem messageを追加して整形したデータセットです。
データの詳細については元データセットのREADMEを参照してください。
ライセンス
MITライセンスの元配布します。
coconot-r1-format-filteredwalton-multimodal-cold-start-r1-format-30k
walton-multimodal-cold-start-r1-format-30k
WaltonFuture/Multimodal-Cold-Start converted to multimodal-open-r1-8k-verified format with filtering
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids:… See the full description on the dataset page: https://huggingface.co/datasets/penfever/walton-multimodal-cold-start-r1-format-30k.tulu_v3.9_wildchat_100k_english-r1-filtered-formatOpenThoughts-114k-r1-formatwildchat-r1-p2-format-filteredopen_math_2_50k_r1-format-filteredwalton-multimodal-cold-start-r1-format
walton-multimodal-cold-start-r1-format
WaltonFuture/Multimodal-Cold-Start converted to multimodal-open-r1-8k-verified format with filtering
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/walton-multimodal-cold-start-r1-format.acecoder-r1-format-filteredAceCoderV2-mini-processed_openrlhf_format_r1rlhf_dataset_20250126_openrlhf_format_hard_r1aime_1983_2023_grok-3-mini-high_traces_32768_r1_formattednuminatmath-r1-format-filtered
