datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
walton-multimodal-cold-start-r1-format
walton-multimodal-cold-start-r1-format
WaltonFuture/Multimodal-Cold-Start converted to multimodal-open-r1-8k-verified format with filtering
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/walton-multimodal-cold-start-r1-format.multimodal-open-r1-8192-filtered-mid-ic
multimodal-open-r1-8192-filtered-mid-ic
Original dataset structure preserved, filtered by token length and image quality
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized input sequences… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/multimodal-open-r1-8192-filtered-mid-ic.s1-vis-mid-resize
s1-vis-mid-resize
Original dataset structure preserved, filtered by token length and image quality
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized input sequences
attention_mask: Attention… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/s1-vis-mid-resize.oumi-walton-exclude-geometry-biologyoumi-walton-exclude-geometry-biology-statisticsoumi-walton-exclude-geometry-biology-statistics-none-of-aboveoumi-web-agents1.1-VLoumi-walton-exclude-geometryoumi-walton-0.7-none-of-above-mathematics-education-0.3-log-weightoumi-walton-include-none-of-aboveoumi-walton-none-of-above-and-up-to-10-for-each-categoryoumi-walton-none-of-above-and-log-weightoumi-walton-0.5-none-of-above-0.5-log-weightlimo-vis-mid-resize
limo-vis-mid-resize
Original dataset structure preserved, filtered by token length and image quality
Dataset Description
This dataset was processed using the data-preproc package for vision-language model training.
Processing Configuration
Base Model: Qwen/Qwen2.5-7B-Instruct
Tokenizer: Qwen/Qwen2.5-7B-Instruct
Sequence Length: 16384
Processing Type: Vision Language (VL)
Dataset Features
input_ids: Tokenized input sequences
attention_mask:… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/limo-vis-mid-resize.
