CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01siyanzhao /Openthoughts_math_30k_opsdtext10K<n<100K10 likes13k downloads7mo agoHugging Face02yunjae-won /openthoughts_math_30k_opsd_splittext10K<n<100K0 likes462 downloads3mo agoHugging Face03HzChen20 /opsd-cl-assets opsd-cl-assets Offline assets for the OPSD continual-learning project on a box without DNS. Built 2026-09-12 on VAST. Pull everything with one command, then run install_on_box.sh. path content size wheelhouse/ 167 wheels resolved from siyan-zhao/OPSD environment.yml for cp310 / manylinux x86_64, plus deepspeed-0.18.2.tar.gz and the lock file 4.7 GB models/Qwen3-1.7B Qwen/Qwen3-1.7B 3.8 GB datasets/Openthoughts_math_30k_opsd siyanzhao/Openthoughts_math_30k_opsd… See the full description on the dataset page: https://huggingface.co/datasets/HzChen20/opsd-cl-assets.text0 likes188 downloads8d agoHugging Face04Camellia86 /Full_Agent_RL_OPSD_with_Just_2_A800stext100K<n<1M0 likes90 downloads23d agoHugging Face05starli-snowflake /swe_opsd_datasettext10K<n<100K0 likes76 downloads3mo agoHugging Face06locaszzz /opsd-tooltext10K<n<100K0 likes76 downloads6d agoHugging Face07LorMolf /SPSD-Variants-opsd SPSD-Variants-opsd Grounded on-policy self-distillation (OPSD) teacher-context dataset over 45 board-game rule variants (5 families × 9: connect4, domineering, simplified_first_attack, simplified_othello, tic_tac_chess), derived from trained MuZero/EfficientZero checkpoints (plan-528 v2). Each row is a decision-state task (a move choice or one of six auxiliary state-QA tasks). The privileged_context is the teacher signal: grounded natural-language reasoning that discovers the… See the full description on the dataset page: https://huggingface.co/datasets/LorMolf/SPSD-Variants-opsd.texttext-generation100K<n<1M0 likes59 downloads23d agoHugging Face08starli-snowflake /scaleswe-opsd-v2-3200-summary Scale-SWE OPSD v2 — 3200 tasks with summary hints The training set used for the Scale-SWE on-policy self-distillation (OPSD) runs. 3200 SWE tasks across 752 repositories, each paired with a reference agent trajectory and a condensed solution hint. Uploaded from /checkpoint/huggingface/datasets/scaleswe_opsd_v2_3200_summary (a datasets.save_to_disk directory), converted to parquet. Row count, ids and field contents verified identical to the source. ⚠️ Contains… See the full description on the dataset page: https://huggingface.co/datasets/starli-snowflake/scaleswe-opsd-v2-3200-summary.texttext-generation1K<n<10K0 likes54 downloads2mo agoHugging Face09mt0rm0 /opsdataDataset Card for Energy — OPEN POWER System Data This dataset was prepared for the OpenHPI course Time Series Analysis and Forecasting and provides electricity consumption, renewable generation, weather, and electricity price data across European countries, including Germany. The data is aggregated from the ENTSO-E Transparency Platform and the Open Power System Data project, spanning over 10 years with 15- and 30-minute time intervals. It is suitable for energy forecasting, renewable… See the full description on the dataset page: https://huggingface.co/datasets/mt0rm0/opsdata.text10M<n<100M1 likes53 downloads1y agoHugging Face10LSW142857 /OPSD-PI-SWE-Gym-512 OPSD-PI SWE-Gym Stage PI 512 Qwen3.5-9B stage-adaptive OPSD-PI 的公开 512-row 数据与 Weak from-scratch 训练包。Public 512-row data and Weak from-scratch training bundle. Files data/train.jsonl: 512 deterministic SWE-Gym rows with Weak, Medium, and Strong PI for EXPLORE, REPRODUCE, DIAGNOSE, EDIT, and VERIFY. data/manifest.json: source selection and integrity metadata. release/OPSD_pi-opsd-pi-weak-from-scratch-20260818.tar.gz: immutable source release containing launchers… See the full description on the dataset page: https://huggingface.co/datasets/LSW142857/OPSD-PI-SWE-Gym-512.texttext-generationn<1K0 likes43 downloads1mo agoHugging Face11hbin0701 /opsd-probe-seed OPSD prefix-continuation probe — seed data Everything needed to reproduce the prefix-continuation probe for OPSD (on-policy self-distillation) on a fresh GPU box, except the base model (Qwen/Qwen3-1.7B, pulled from the Hub at setup) and the code repo (hbin0701/OPSD). These artifacts live outside git because the training/eval output directory is .gitignored. What the probe answers Fitting p' = p + λ·(1[mode correct] − p) + γ against a properly sampled 64-shot… See the full description on the dataset page: https://huggingface.co/datasets/hbin0701/opsd-probe-seed.tabulartext-generationn<1K0 likes36 downloads1mo agoHugging Face12hmnsyrd /ops-datatext1B<n<10B0 likes34 downloads29d agoHugging Face13SeongryongJung /opsd-plain-4b-rollouts opsd-plain-4b-rollouts This dataset contains rollout generations collected during training. Source experiment method: opsd-plain model_size: 4b experiment_dir: /home/irteam/outputs/opsd_plain_4b Format Each row contains: step sample_index prompt completion method model_size source_file Viewer structure all: all rollout rows together step_<N>: only one rollout step, easier to inspect in the dataset viewer Notes… See the full description on the dataset page: https://huggingface.co/datasets/SeongryongJung/opsd-plain-4b-rollouts.tabulartext-generationn<1K0 likes26 downloads4mo agoHugging Face14Keven16 /OPSD-Example-Datatext10K<n<100K0 likes25 downloads6mo agoHugging Face15opsd-genrm /dedup_filtered_HS3 Dedup-Filtered HelpSteer3 (+ rubrics on train) A cleaned version of nvidia/HelpSteer3 (preference config) with three modifications: Filter out domain == "multilingual" and overall_preference == 0 (tie) rows. Within-split dedup by content hash (sha1(context, response1, response2)), keeping the first occurrence. Upstream HS3 has ~35% byte-identical duplicate rows after the filter step. Cross-split dedup: drop validation rows whose content hash also appears in train. Upstream HS3… See the full description on the dataset page: https://huggingface.co/datasets/opsd-genrm/dedup_filtered_HS3.texttext-classification10K<n<100K0 likes16 downloads5mo agoHugging Face16SeongryongJung /opsd-plain-8b-rollouts opsd-plain-8b-rollouts This dataset contains rollout generations collected during training. Source experiment method: opsd-plain model_size: 8b experiment_dir: /home/irteam/outputs/opsd_plain_8b Format Each row contains: step sample_index prompt completion method model_size source_file Viewer structure all: all rollout rows together step_<N>: only one rollout step, easier to inspect in the dataset viewer Notes… See the full description on the dataset page: https://huggingface.co/datasets/SeongryongJung/opsd-plain-8b-rollouts.tabulartext-generationn<1K0 likes16 downloads4mo agoHugging Face17violetxi /olympiad_physics_stage1_qwen8b_opsdtext10K<n<100K0 likes12 downloads4mo agoHugging Face18dvyomkesh /opsd_edge_dataset OPSD Edge Dataset (Paper-Faithful) Edge prompts for On-Policy Self-Distillation training on Nemotron-3-Nano. Paper Reference "Self-Distilled Reasoner: On-Policy Self-Distillation for LLMs" Paper: arXiv:2601.18734 Code: github.com/siyan-zhao/OPSD Dataset Description This dataset contains 1,782 "edge" prompts where the 0.83 Nemotron adapter achieves 25-75% pass rate (uncertain cases ideal for learning). Paper-Faithful Format From Figure 2 of the OPSD… See the full description on the dataset page: https://huggingface.co/datasets/dvyomkesh/opsd_edge_dataset.tabulartext-generation1K<n<10K0 likes11 downloads4mo agoHugging Face19violetxi /aime25_stage1_qwen8b_opsdtext1K<n<10K0 likes10 downloads4mo agoHugging Face20cometadata /opsdc-training-datatext1K<n<10K0 likes9 downloads6mo agoHugging Face21ayuzhenyu /Openthoughts_math_30k_opsdtext10K<n<100K0 likes8 downloads1mo agoHugging Face22violetxi /aime26_stage1_qwen8b_opsdtext1K<n<10K1 likes7 downloads4mo agoHugging Face23violetxi /lcb_v5_stage1_qwen8b_opsdtext1K<n<10K0 likes6 downloads4mo agoHugging Face24violetxi /gpqa_diamond_stage1_qwen8b_opsdtext10K<n<100K0 likes6 downloads4mo agoHugging Face25violetxi /imo-answerbench_stage1_qwen8b_opsdtext10K<n<100K0 likes5 downloads4mo agoHugging Face26violetxi /hmmt-nov-2025_stage1_qwen8b_opsdtext1K<n<10K0 likes4 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.