CoolFace
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RabotniKuma /Fast-Math-R1-GRPOThis repository contains the second-stage GRPO dataset for the paper A Practical Two-Stage Recipe for Mathematical LLMs: Maximizing Accuracy with SFT and Efficiency with Reinforcement Learning. This dataset is crucial for the second stage of the training recipe, aiming to improve token efficiency while preserving peak mathematical reasoning performance in Large Language Models (LLMs) through Reinforcement Learning from online inference (GRPO). We extracted the answers from the 2nd stage SFT… See the full description on the dataset page: https://huggingface.co/datasets/RabotniKuma/Fast-Math-R1-GRPO.texttext-generation1K<n<10K2 likes60 downloads1y agoHugging Face02predibase /wordle-grpotextn<1K4 likes27 downloads1y agoHugging Face03Anna4242 /grpo-training-plotstabular1K<n<10K0 likes21 downloads10mo agoHugging Face04AndyZhong1953 /grpo_auto_evaltabular10K<n<100K0 likes19 downloads8mo agoHugging Face05saracandu /eureka-rebus-grpotext10K<n<100K0 likes14 downloads1y agoHugging Face06sdzt /forensics-grpo-data forensics-grpo-data Generated-video dataset + annotations used to train sdzt/forensics-grpo. 📂 Repository layout forensics-grpo-data/ ├── video/ # 5,388 .mp4 clips, packed as one .tar per generator │ ├── 01_vidu.tar # 9.7 GB — Vidu │ ├── 02_wan.tar # 28 GB — Wan │ ├── 03_fcvg.tar # 27 GB — FCVG │ ├── 04_scifi.tar # 34 GB — SciFi │ ├── 05_ltx.tar # 6.3 GB — LTX │… See the full description on the dataset page: https://huggingface.co/datasets/sdzt/forensics-grpo-data.textvideo-classification1K<n<10K0 likes10 downloads4mo agoHugging Face07agentic-moral-alignment /qwen35-9b-grpo-unsloth-ut-tft-1000eptext1K<n<10K0 likes8 downloads6mo agoHugging Face08CodCodingCode /grpo-usagetextn<1K0 likes7 downloads1y agoHugging Face09agentic-moral-alignment /qwen35-9b-grpo-unsloth-game-tft-1000eptext1K<n<10K0 likes5 downloads6mo agoHugging Face10nolangclem /bigmath-grpo-rolloutstabular100K<n<1M0 likes4 downloads7mo agoHugging Face11anson1788 /GRPOtestingtext1K<n<10K0 likes1 downloads2y agoHugging Face12Rupesh2 /Evaluation_GRPOgatedtabulartext-generation1K<n<10K0 likes1 downloads2y agoHugging Face13akbarsigit /llama3.1-grpo-r128-a256_log_20250525_123129tabularn<1K0 likes1 downloads1y agoHugging Face14akbarsigit /llama3.1-grpo-r256-a512_log_20250526_114710tabularn<1K0 likes1 downloads1y agoHugging Face15nolangclem /bigmath-grpo-rollouts-qwen25-3btabular100K<n<1M0 likes1 downloads7mo agoHugging Face16akbarsigit /llama3.1-grpo-r256-a512-base_log_20250526_165221tabularn<1K0 likes1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.