oumi
Datasets
All datasets matching “oumi”MetaMathQA-R1
oumi-ai/MetaMathQA-R1
MetaMathQA-R1 is a text dataset designed to train Conversational Language Models with DeepSeek-R1 level reasoning.
Prompts were augmented from GSM8K and MATH training sets with responses directly from DeepSeek-R1.
MetaMathQA-R1 was used to train MiniMath-R1-1.5B, which achieves 44.4% accuracy on MMLU-Pro-Math, the highest of any model with <=1.5B parameters.
Curated by: Oumi AI using Oumi inference on Parasail
Language(s) (NLP): English
License:… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/MetaMathQA-R1.lmsys_chat_1m_clean_R1
oumi-ai/lmsys_chat_1m_clean_R1
lmsys_chat_1m_clean_R1 is a text dataset designed to train Conversational Language Models with DeepSeek-R1 level reasoning.
Prompts were pulled from LMSYS and filtered to lmsys_chat_1m_clean, and responses were taken from DeepSeek-R1 without additional filters present.
We release lmsys_chat_1m_clean_R1 to help enable the community to develop the best fully open reasoning model!
lmsys_chat_1m_clean queries with responses generated from… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/lmsys_chat_1m_clean_R1.oumi-letter-countbanking77-oumi-quickstartMM-MathInstruct-to-r1-format-filtered
MM-MathInstruct-to-r1-format-filtered
MM-MathInstruct dataset transformed to R1 format and filtered by token length and image quality
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized input… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/MM-MathInstruct-to-r1-format-filtered.oumi-ai_lmsys_chat_1m_clean_R1-1k-think-1k-response-ShareGPT
