CoolFace
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01allenai /Dolci-RL-Zero-Code-7B Dolci RL-Zero Code Dolci RL-Zero Code is a dataset of 13.3k coding questions and answers for RLVR training of Olmo 3 7B RL-Zero Code This dataset was collected from the code subset of Dolci Think SFT 7B, see the Olmo 3 paper for details. Downloading You can download and load this data using HuggingFace's datasets library with the following code: from datasets import load_dataset dataset = load_dataset("allenai/Dolci-RL-Zero-Code-7B", split="train",) Licensing… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-RL-Zero-Code-7B.textreinforcement-learning10K<n<100K10 likes209 downloads9mo agoHugging Face02allenai /Dolci-RL-Zero-Math-7B Dolci RL-Zero Math Dolci RL-Zero Math is a dataset of 13.3k math questions and answers for RLVR training of Olmo 3 RL-Zero Math 7B This dataset was collected from a subset of DAPO Math and Klear Reasoner Math, see the Olmo 3 paper for details. Downloading You can download and load this data using HuggingFace's datasets library with the following code: from datasets import load_dataset dataset = load_dataset("allenai/Dolci-RL-Zero-Math-7B", split="train",)… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-RL-Zero-Math-7B.textreinforcement-learning10K<n<100K10 likes189 downloads9mo agoHugging Face03allenai /Dolci-RL-Zero-IF-7B Dolci RL-Zero IF Dolci RL-Zero IF is a dataset of 13.3k instruction-following prompts and answers for RLVR training of Olmo 3 7B RL-Zero IF This dataset was collected from the instruction-following subset of Dolci Think SFT 7B, see the Olmo 3 paper for details. Downloading You can download and load this data using HuggingFace's datasets library with the following code: from datasets import load_dataset dataset = load_dataset("allenai/Dolci-RL-Zero-IF-7B", split="train",)… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-RL-Zero-IF-7B.textreinforcement-learning10K<n<100K5 likes141 downloads9mo agoHugging Face04allenai /Dolci-RL-Zero-Mix-7B Dolci RL-Zero Mix Dolci RL-Zero Mix is a dataset of 39.9k math, code, and instruction-following prompts and answers for RLVR training of Olmo 3 7B RL-Zero Mix This dataset is mix of Dolci RLZero Math 7B, Dolci RLZero Code 7B, Dolci RLZero IF 7B, and Dolci RL Zero General 7B, see the Olmo 3 paper for details. Downloading You can download and load this data using HuggingFace's datasets library with the following code:from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-RL-Zero-Mix-7B.text10K<n<100K0 likes125 downloads10mo agoHugging Face05THU-KEG /LongWriter-Zero-RLData LongWriter-Zero RL Data 🤗 [Model] • 📃 [Paper] • 💾 [Dataset Card] LongWriter-Zero RL Data is designed for ultra-long text generation via reinforcement learning. The dataset consists of conversational queries paired with length-range tags, which specify the desired output span (measured in words or Chinese characters). These annotations are used to train the LongWriter-Zero model, enabling it to consistently generate passages exceeding 10,000 words. PS: We also included some… See the full description on the dataset page: https://huggingface.co/datasets/THU-KEG/LongWriter-Zero-RLData.texttext-generation1K<n<10K24 likes112 downloads1y agoHugging Face06mlfoundations-cua-dev /easyr1-103k-4MP-jedi-ui-vision-gta1-data-sampling-stage-three-temp-1_7-RL-zero-correct-to-0.2image10K<n<100K0 likes112 downloads1y agoHugging Face07allenai /Dolci-RL-Zero-General-7B Dolci-RL-Zero-General-7B Dataset Summary Dolci-RL-Zero-General-7B is the reinforcement learning dataset used to train the Olmo3-RL-Zero-7B-General model.It contains 12,841 general chat prompts sampled from the larger Dolci-Think-RL mixture. The reward was dervied by using an LM judge. Downloading You can download and load this data using HuggingFace's datasets library with the following code: from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-RL-Zero-General-7B.textreinforcement-learning10K<n<100K4 likes64 downloads9mo agoHugging Face08mlfoundations-cua-dev /easyr1-103k-4MP-jedi-ui-vision-gta1-data-sampling-stage-two-temp-1_1-RL-zero-correct-to-0.3image10K<n<100K0 likes54 downloads1y agoHugging Face09mlfoundations-cua-dev /easyr1-103k-4MP-jedi-ui-vision-gta1-data-sampling-stage-two-temp-1_1-RL-zero-correct-to-0.2image10K<n<100K0 likes35 downloads1y agoHugging Face10heegyu /Dolci-RL-Zero-Code-7Btext10K<n<100K0 likes29 downloads9mo agoHugging Face11mlfoundations-cua-dev /easyr1-103k-4MP-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktop-gta-0p0-zeroimage1K<n<10K0 likes26 downloads1y agoHugging Face12HFXM /response-data-MiMo-7B-RL-Zerotext10K<n<100K0 likes26 downloads7mo agoHugging Face13heegyu /Dolci-RL-Zero-Code-7B-refinedtext10K<n<100K0 likes24 downloads8mo agoHugging Face14waleko /Dolci-RL-Zero-Math-7B-solvedtext1K<n<10K0 likes23 downloads7mo agoHugging Face15mlfoundations-cua-dev /easyr1-103k-4MP-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktop-0p2-zeroimage1K<n<10K0 likes21 downloads1y agoHugging Face16aochongoliverli /Qwen2.5-1.5B-deepmath-level1-3-rl-zero-4rollout-4096max-len-rolloutstext1K<n<10K0 likes17 downloads1y agoHugging Face17mlfoundations-cua-dev /easyr1-103k-4MP-stage-three-temp-1-7-RL-zero-to-0.2-noise-DONT-TRAIN-ONimage1K<n<10K0 likes14 downloads1y agoHugging Face18Slim205 /lean_workbook_RL_no_zero_examples_1000tabular1K<n<10K0 likes9 downloads1y agoHugging Face19mlfoundations-cua-dev /easyr1-103k-4MP-stage-three-temp-1-7-RL-zero-to-0.2-no-pixmo-uground-seeclickimage1K<n<10K0 likes9 downloads1y agoHugging Face20heegyu /Dolci-RL-Zero-IF-7Btext10K<n<100K0 likes9 downloads9mo agoHugging Face21Slim205 /lean_workbook_RL_no_zero_examples_2000tabular10K<n<100K0 likes8 downloads1y agoHugging Face22pittawat /Dolci-RL-Zero-Math-7B-classifiedtext10K<n<100K0 likes8 downloads10mo agoHugging Face23Slim205 /lean_workbook_RL_no_zero_examplestabular1K<n<10K0 likes7 downloads1y agoHugging Face24mlfoundations-cua-dev /easyr1-103k-4MP-stage-three-temp-1_7-RL-only-ui-vision-jedi-show-ui-desktop-gta-0p2-zeroimage1K<n<10K0 likes5 downloads1y agoHugging Face25mlfoundations-cua-dev /easyr1-103k-4MP-stage-three-temp-1_7-RL-ui-vision-jedi-show-ui-desktop-gta-0p2-zero-dense-rewardimage1K<n<10K0 likes4 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.