datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ru-thinking-reasoning-r1-v2-dedupedru-thinking-reasoning-r1Combined dataset of mostly Russian thinking/reasoning/reflection dialogs in form of conversation suitable for LLM fine-tuning scenarios. All responses are mapped to same format.
The format of reasoning in most cases is:
<think>
Reasoning...
</think>
Response
For reflection dataset - there can be also <reflection> tags inside <think>.
Common system prompt for think:
Ты полезный ассистент. Отвечай на вопросы, сохраняя следующую структуру: <think> Твои мысли и рассуждения </think>
Твой конечный… See the full description on the dataset page: https://huggingface.co/datasets/ZeroAgency/ru-thinking-reasoning-r1.ru-thinking-reasoning-r1-v2This is a modified version of ZeroAgency/ru-thinking-reasoning-r1 with addition of Egor-AI/CoT-XLang dataset.
Combined dataset of mostly Russian thinking/reasoning/reflection dialogs in form of conversation suitable for LLM fine-tuning scenarios. All responses are mapped to same format.
The format of reasoning in most cases is:
<think>
Reasoning...
</think>
Response
For reflection dataset - there can be also <reflection> tags inside <think>.
Common system prompt for think:
Ты полезный… See the full description on the dataset page: https://huggingface.co/datasets/ZeroAgency/ru-thinking-reasoning-r1-v2.Table-R1-Zero-Datasettrn-r1-zero-tagsru-thinking-reasoning-r1-dedupedStrikeGPT-R1-Zero-Test-CoT_DataNetwork Security Reasoning Dataset(test)
Obtained through distillation of DeepSeek-R1
think to CoT code:https://github.com/Bouquets-ai/Data-Processing/blob/main/think%20to%20CoT.py
R1-Zero-GRPO-7500table-r1-zero-verl
Table-R1-Zero (VERL Format)
This dataset contains 69,265 table reasoning problems from the Table-R1-Zero-Dataset, converted to VERL (Volcano Engine Reinforcement Learning) format for reinforcement learning training workflows.
Source: Table-R1/Table-R1-Zero-Dataset
License: Apache 2.0
Note: System prompts have been removed from all examples for better compatibility with other VERL datasets. The dataset now contains only user messages with table reasoning problems. Ground truth… See the full description on the dataset page: https://huggingface.co/datasets/sungyub/table-r1-zero-verl.R1-Zero-GRPO-750square-d1-dagger-sobol-v1-hard-zerofive-r1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "panda",
"total_episodes": 200,
"total_frames": 45738,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 20,
"splits": {
"train": "0:200"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankile/square-d1-dagger-sobol-v1-hard-zerofive-r1.R1-Zero-GRPO-1500UniVG-R1-zeroshotDeepSeek-R1-Zero-best_of_n-VLLM-Skywork-o1-Open-PRM-Qwen-2.5-7B-completions
