CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01allenai /rlvr-code-data-python-r1-format-filteredtabular10K<n<100K4 likes389 downloads1y agoHugging Face02koyena /Magpie-Reasoning-V2-250K-CoT-Deepseek-R1-Llama-70B-formattedtext100K<n<1M0 likes219 downloads1y agoHugging Face03jonathanyin /aime_1983_2023_grok-3-mini-high_traces_r1_formattedtabularn<1K0 likes134 downloads1y agoHugging Face04allenai /wildjailbreak-r1-v2-format-filteredtext10K<n<100K4 likes123 downloads1y agoHugging Face05xDAN-Vision /MMMU-LLM-R1-format MMMU-LLM-R1 Reformatted Dataset imagen<1K0 likes74 downloads2y agoHugging Face06Aratako /Synthetic-Japanese-Roleplay-SFW-DeepSeek-R1-0528-10k-formatted Synthetic-Japanese-Roleplay-SFW-DeepSeek-R1-0528-10k-formatted 概要 deepseek-ai/DeepSeek-R1-0528を用いて作成した日本語ロールプレイデータセットであるAratako/Synthetic-Japanese-Roleplay-SFW-DeepSeek-R1-0528-10kにsystem messageを追加して整形したデータセットです。 データの詳細については元データセットのREADMEを参照してください。 ライセンス MITライセンスの元配布します。 texttext-generation10K<n<100K0 likes74 downloads1y agoHugging Face07saumyamalik /rlvr-code-data-python-r1-format-filteredtabular10K<n<100K0 likes65 downloads1y agoHugging Face08saumyamalik /wildchat-r1-p2-format-filteredtabular10K<n<100K0 likes59 downloads1y agoHugging Face09di-zhang-fdu /llava-cot-100k-r1-format llava-cot-100k-r1-format: A dataset for Vision Reasoning GRPO Training Images Images data can be access from https://huggingface.co/datasets/Xkev/LLaVA-CoT-100k SFT dataset https://huggingface.co/datasets/di-zhang-fdu/R1-Vision-Reasoning-Instructions Citations @misc {di_zhang_2025, author = { {Di Zhang} }, title = { llava-cot-100k-r1-format (Revision 87d607e) }, year = 2025, url = {… See the full description on the dataset page: https://huggingface.co/datasets/di-zhang-fdu/llava-cot-100k-r1-format.text100K<n<1M2 likes57 downloads1y agoHugging Face10allenai /the-algorithm-python-r1-format-filteredtextn<1K1 likes56 downloads1y agoHugging Face11oumi-ai /MM-MathInstruct-to-r1-format-filtered MM-MathInstruct-to-r1-format-filtered MM-MathInstruct dataset transformed to R1 format and filtered by token length and image quality Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids: Tokenized input… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/MM-MathInstruct-to-r1-format-filtered.text10K<n<100K0 likes53 downloads1y agoHugging Face12saumyamalik /wildchat-r1-p2-filtered-formattabular10K<n<100K0 likes52 downloads1y agoHugging Face13saumyamalik /persona-precise-if-r1-merged-format-filteredtext100K<n<1M0 likes46 downloads1y agoHugging Face14DylanDDeng /default-open-r1-math-90k-formatThis dataset, derived from the default version of OpenR1-Math-220K, has been reformatted and organized for improved usability and model training. The following data processing and quality filtering measures were implemented: Data Processing and Quality Filtering Methodology Structural Integrity Validation: Ensures data consistency by verifying equal lengths across correctness_math_verify, is_reasoning_complete, and generations lists. Confirms the presence of all required fields within each… See the full description on the dataset page: https://huggingface.co/datasets/DylanDDeng/default-open-r1-math-90k-format.textquestion-answering10K<n<100K1 likes44 downloads2y agoHugging Face15jonathanyin /aime_1983_2023_grok-3-mini-high_traces_16392_r1_formattedtabularn<1K0 likes43 downloads1y agoHugging Face16saumyamalik /aya-100k-r1-format-filteredtext10K<n<100K0 likes39 downloads1y agoHugging Face17saumyamalik /tulu_v3.9_wildchat_100k_english-r1-format-filteredtext10K<n<100K0 likes37 downloads1y agoHugging Face18Aratako /Synthetic-Japanese-Roleplay-NSFW-DeepSeek-R1-0528-10k-formatted Synthetic-Japanese-Roleplay-NSFW-DeepSeek-R1-0528-10k-formatted 概要 deepseek-ai/DeepSeek-R1-0528を用いて作成した日本語ロールプレイデータセットであるAratako/Synthetic-Japanese-Roleplay-NSFW-DeepSeek-R1-0528-10kにsystem messageを追加して整形したデータセットです。 データの詳細については元データセットのREADMEを参照してください。 ライセンス MITライセンスの元配布します。 texttext-generation10K<n<100K0 likes35 downloads1y agoHugging Face19saumyamalik /coconot-r1-format-filteredtext10K<n<100K0 likes35 downloads1y agoHugging Face20penfever /walton-multimodal-cold-start-r1-format-30k walton-multimodal-cold-start-r1-format-30k WaltonFuture/Multimodal-Cold-Start converted to multimodal-open-r1-8k-verified format with filtering Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids:… See the full description on the dataset page: https://huggingface.co/datasets/penfever/walton-multimodal-cold-start-r1-format-30k.image10K<n<100K1 likes34 downloads1y agoHugging Face21saumyamalik /tulu_v3.9_wildchat_100k_english-r1-filtered-formattext10K<n<100K0 likes32 downloads1y agoHugging Face22di-zhang-fdu /OpenThoughts-114k-r1-formattext100K<n<1M0 likes31 downloads2y agoHugging Face23allenai /wildchat-r1-p2-format-filteredtabular10K<n<100K2 likes29 downloads1y agoHugging Face24allenai /open_math_2_50k_r1-format-filteredtext10K<n<100K1 likes27 downloads1y agoHugging Face25oumi-ai /walton-multimodal-cold-start-r1-format walton-multimodal-cold-start-r1-format WaltonFuture/Multimodal-Cold-Start converted to multimodal-open-r1-8k-verified format with filtering Dataset Description This dataset was processed using the data-preproc package for vision-language model training. Processing Configuration Base Model: Qwen/Qwen2.5-7B-Instruct Tokenizer: Qwen/Qwen2.5-7B-Instruct Sequence Length: 16384 Processing Type: Vision Language (VL) Dataset Features input_ids: Tokenized… See the full description on the dataset page: https://huggingface.co/datasets/oumi-ai/walton-multimodal-cold-start-r1-format.image10K<n<100K1 likes26 downloads1y agoHugging Face26saumyamalik /acecoder-r1-format-filteredtabular10K<n<100K0 likes25 downloads1y agoHugging Face27CodeDPO /AceCoderV2-mini-processed_openrlhf_format_r1text10K<n<100K0 likes24 downloads2y agoHugging Face28CodeDPO /rlhf_dataset_20250126_openrlhf_format_hard_r1text10K<n<100K0 likes23 downloads2y agoHugging Face29jonathanyin /aime_1983_2023_grok-3-mini-high_traces_32768_r1_formattedtabularn<1K0 likes23 downloads1y agoHugging Face30saumyamalik /numinatmath-r1-format-filteredtext10K<n<100K0 likes23 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.