CoolFace
19 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01reasoning-cues /rollouts-olmo7b-cue-search rollouts-olmo7b-cue-search Model: allenai/Olmo-3-1025-7B (snapshot a81bae42). Tokenizer: allenai/Olmo-3-1025-7B (snapshot a81bae42). Protocol: RL-Zero prompt, MATH-500 x 4 rollouts, budget 31,744, T 0.6, top-p 0.95, seed 20260819 (depth-2 exhaustive and n-gram chain: seed 20260821); the top-20 beam nominee screen, ten random-opener arms, every depth-2 opener (84 shards, arm names unique across shards) and the n-gram chain arms. Rollouts generated on the CSAIL cluster for the… See the full description on the dataset page: https://huggingface.co/datasets/reasoning-cues/rollouts-olmo7b-cue-search.tabular100K<n<1M0 likes883 downloads10d agoHugging Face02SeanWang0027 /polaris_rose_rollouts_olmo3-7b_from_qwen3-30b-a3b_cutoff4096_240steps Cross-tokenizer ROSE rollouts — Olmo-3-7B-Think-SFT ← Qwen3-30B-A3B-Thinking-2507 Every assembled row of a complete 240-step online-ROSE run: 61,440 rows, the teacher's actual continuation for each, and the token accounting behind it. The student writes a 4096-token prefix in its own vocabulary (100278). That prefix is decoded to text, the teacher is shown it under its own chat template, and the teacher's reply comes back as text and is tokenised into the student's vocabulary.… See the full description on the dataset page: https://huggingface.co/datasets/SeanWang0027/polaris_rose_rollouts_olmo3-7b_from_qwen3-30b-a3b_cutoff4096_240steps.tabulartext-generation10K<n<100K0 likes54 downloads24d agoHugging Face03Mihara-bot /olmo-igsm-arith OLMo iGSM-Easy Arithmetic This repository contains a frozen, evaluation-only release of the synthetic mod-7 arithmetic task called iGSM-Easy Arithmetic in the accompanying OLMo evaluation code. It contains 750 examples: 250 examples at each target depth 2, 3, and 4. This is an i-GSM-style task variant, not a claim to be an official release of another dataset named iGSM. The olmo-igsm-arith name is used to make the implementation provenance explicit. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Mihara-bot/olmo-igsm-arith.tabularquestion-answeringn<1K0 likes36 downloads2mo agoHugging Face04yakazimir /ultrafeedback_olmo1b_reftabular10K<n<100K0 likes34 downloads2y agoHugging Face05open-llm-leaderboard /allenai__OLMo-7B-hf-detailsgated Dataset Card for Evaluation run of allenai/OLMo-7B-hf Dataset automatically created during the evaluation run of model allenai/OLMo-7B-hf The dataset is composed of 39 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMo-7B-hf-details.tabular10K<n<100K0 likes28 downloads2y agoHugging Face06open-llm-leaderboard /allenai__OLMo-1.7-7B-hf-detailsgated Dataset Card for Evaluation run of allenai/OLMo-1.7-7B-hf Dataset automatically created during the evaluation run of model allenai/OLMo-1.7-7B-hf The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMo-1.7-7B-hf-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face07open-llm-leaderboard /allenai__OLMo-7B-Instruct-hf-detailsgated Dataset Card for Evaluation run of allenai/OLMo-7B-Instruct-hf Dataset automatically created during the evaluation run of model allenai/OLMo-7B-Instruct-hf The dataset is composed of 39 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMo-7B-Instruct-hf-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face08open-llm-leaderboard /allenai__OLMo-1B-hf-detailsgated Dataset Card for Evaluation run of allenai/OLMo-1B-hf Dataset automatically created during the evaluation run of model allenai/OLMo-1B-hf The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMo-1B-hf-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face09open-llm-leaderboard /allenai__OLMoE-1B-7B-0924-Instruct-detailsgated Dataset Card for Evaluation run of allenai/OLMoE-1B-7B-0924-Instruct Dataset automatically created during the evaluation run of model allenai/OLMoE-1B-7B-0924-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMoE-1B-7B-0924-Instruct-details.tabular10K<n<100K0 likes22 downloads2y agoHugging Face10open-llm-leaderboard /allenai__OLMo-2-1124-7B-Instruct-detailsgated Dataset Card for Evaluation run of allenai/OLMo-2-1124-7B-Instruct Dataset automatically created during the evaluation run of model allenai/OLMo-2-1124-7B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMo-2-1124-7B-Instruct-details.tabular10K<n<100K0 likes22 downloads2y agoHugging Face11open-llm-leaderboard /allenai__OLMoE-1B-7B-0125-Instruct-detailsgated Dataset Card for Evaluation run of allenai/OLMoE-1B-7B-0125-Instruct Dataset automatically created during the evaluation run of model allenai/OLMoE-1B-7B-0125-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMoE-1B-7B-0125-Instruct-details.tabular10K<n<100K0 likes22 downloads2y agoHugging Face12open-llm-leaderboard /allenai__OLMoE-1B-7B-0924-detailsgated Dataset Card for Evaluation run of allenai/OLMoE-1B-7B-0924 Dataset automatically created during the evaluation run of model allenai/OLMoE-1B-7B-0924 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allenai__OLMoE-1B-7B-0924-details.tabular10K<n<100K0 likes20 downloads2y agoHugging Face13Nwna /olmo3-190m-zh-v2-base-data OLMo3-190M-zh-v2 Base Tokenized Data 这是 OLMo3-190M 中文 v2 base pretrain 使用的正式 tokenized 数据。 本仓库只发布 packed token id,不发布 raw text shards。这样可以让训练复现直接读取 tokenized.bin,同时避免原始语料再分发带来的体积和授权边界问题。 文件说明 tokenized.bin # uint16 一维 token 流 meta.json # 数据构建、混合比例、token 统计等元数据 必须匹配的 Tokenizer 本数据必须使用下面这个 tokenizer 解码和训练: Nwna/olmo3-190m-zh-v2-tokenizer 不要用其他 tokenizer 读取这份 tokenized.bin。同一个数字 token id 在不同 tokenizer 里含义不同,混用会导致训练目标错位,即使 loss 下降也可能训练出坏模型。 数据摘要… See the full description on the dataset page: https://huggingface.co/datasets/Nwna/olmo3-190m-zh-v2-base-data.tabulartext-generationn<1K0 likes10 downloads5mo agoHugging Face14CL-From-Nothing /kukurasu-allenai_OLMo-3-7B-Thinktabular10K<n<100K0 likes9 downloads6mo agoHugging Face15NewEden /rollouts-olmo3tabular10K<n<100K0 likes8 downloads9mo agoHugging Face16CL-From-Nothing /minesweeper-allenai_OLMo-3-7B-Think-continued-by-Qwen_Qwen3-4B-trunc4096tabular10K<n<100K0 likes7 downloads6mo agoHugging Face17CL-From-Nothing /minesweeper-allenai_OLMo-3-7B-Think-continued-by-Qwen_Qwen3-4B-Thinking-trunc4096-resp16384tabular10K<n<100K0 likes6 downloads6mo agoHugging Face18CL-From-Nothing /minesweeper-allenai_OLMo-3-7B-Thinktabular10K<n<100K0 likes4 downloads6mo agoHugging Face19CL-From-Nothing /kukurasu-allenai_OLMo-3-7B-Think-continued-by-Qwen_Qwen3-4B-trunc4096tabular10K<n<100K0 likes4 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.