CoolFace
15 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01oumi-ai /walton-multimodal-cold-start-r1-format walton-multimodal-cold-start-r1-format WaltonFuture/Multimodal-Cold-Start converted to multimodal-open-r1-8k-verified format with filtering Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids: Tokenized… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/walton-multimodal-cold-start-r1-format.image10K<n<100K1 likes26 downloads1y agoHugging Face02oumi-ai /multimodal-open-r1-8192-filtered-mid-ic multimodal-open-r1-8192-filtered-mid-ic Original dataset structure preserved, filtered by token length and image quality Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids: Tokenized input sequences… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/multimodal-open-r1-8192-filtered-mid-ic.image1K<n<10K0 likes24 downloads1y agoHugging Face03oumi-ai /s1-vis-mid-resize s1-vis-mid-resize Original dataset structure preserved, filtered by token length and image quality Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids: Tokenized input sequences attention_mask: Attention… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/s1-vis-mid-resize.imagen<1K0 likes23 downloads1y agoHugging Face04yosubshin /oumi-walton-exclude-geometry-biologyimage1K<n<10K0 likes17 downloads10mo agoHugging Face05yosubshin /oumi-walton-exclude-geometry-biology-statisticsimage1K<n<10K1 likes14 downloads10mo agoHugging Face06yosubshin /oumi-walton-exclude-geometry-biology-statistics-none-of-aboveimage1K<n<10K0 likes13 downloads10mo agoHugging Face07shanghong /oumi-web-agentimage1K<n<10K0 likes11 downloads1y agoHugging Face08oumi-ai /s1.1-VLimage1K<n<10K0 likes8 downloads1y agoHugging Face09yosubshin /oumi-walton-exclude-geometryimage1K<n<10K0 likes7 downloads10mo agoHugging Face10yosubshin /oumi-walton-0.7-none-of-above-mathematics-education-0.3-log-weightimage1K<n<10K0 likes5 downloads10mo agoHugging Face11yosubshin /oumi-walton-include-none-of-aboveimage1K<n<10K0 likes4 downloads10mo agoHugging Face12yosubshin /oumi-walton-none-of-above-and-up-to-10-for-each-categoryimagen<1K0 likes3 downloads10mo agoHugging Face13yosubshin /oumi-walton-none-of-above-and-log-weightimage1K<n<10K0 likes3 downloads10mo agoHugging Face14yosubshin /oumi-walton-0.5-none-of-above-0.5-log-weightimage1K<n<10K0 likes3 downloads10mo agoHugging Face15oumi-ai /limo-vis-mid-resize limo-vis-mid-resize Original dataset structure preserved, filtered by token length and image quality Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids: Tokenized input sequences attention_mask:… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/limo-vis-mid-resize.imagen<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.