CoolFace
23 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nvidia /Nemotron-SFT-ARC-AGI-v1 Dataset Description: Nemotron-SFT-ARC-AGI-v1 is a supervised fine-tuning (SFT) dataset of multi-turn agentic reasoning traces produced by open-weight large language models attempting to solve ARC-AGI visual-reasoning puzzles. Each ARC puzzle (a set of (input grid, output grid) demonstration pairs plus one or more test inputs, where grids are 2D integer arrays representing colors) is formatted as a text prompt and given to an agent powered by one of nine open-weight reasoning… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-SFT-ARC-AGI-v1.texttext-generation100K<n<1M22 likes1.9k downloads4mo agoHugging Face02zhmz90 /arc-agi-2text1K<n<10K1 likes1.7k downloads1y agoHugging Face03AgentNativeResearchLab /arc-agi3-kimi-k2.7-ar25 ARC-AGI-3 ar25 — Agent Trajectories (kimi-k2.7) Gameplay trajectories from the harness×model pair kimi-k2.7 playing the ARC-AGI-3 game ar25, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game played by… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-kimi-k2.7-ar25.tabularreinforcement-learningn<1K0 likes1.1k downloads26d agoHugging Face04AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-tr87 ARC-AGI-3 tr87 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game tr87, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-tr87.tabularreinforcement-learningn<1K0 likes702 downloads2mo agoHugging Face05AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-g50t ARC-AGI-3 g50t — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game g50t, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-g50t.tabularreinforcement-learningn<1K0 likes621 downloads2mo agoHugging Face06nvidia /Nemotron-RL-ARC-AGI-v1 Dataset Description: Nemotron-RL-ARC-AGI-v1 is a reinforcement-learning (RL) gym environment dataset of single-step ARC-AGI puzzle prompts intended for RL post-training of large language models. Each row corresponds to one ARC puzzle (a set of (input grid, output grid) demonstration pairs plus a single test input grid) rendered as a text prompt; reward is binary (1.0 / 0.0) determined by exact-match comparison against the ground-truth output grid. No LLM judge is used, no… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-ARC-AGI-v1.texttext-generation10K<n<100K11 likes590 downloads4mo agoHugging Face07AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-su15 ARC-AGI-3 su15 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game su15, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-su15.tabularreinforcement-learningn<1K0 likes402 downloads2mo agoHugging Face08cveinnt /kepler-arc-agi-3-traces Kepler 1.0 ARC-AGI-3 trace corpus Run artifacts from Kepler 1.0, an open-source agent harness for the 25 public ARC-AGI-3 games. A stock CLI coding agent encodes its theory of each game as an executable world_model.py, certifies it against the full recorded interaction history, plans inside the certified model, and acts through a guarded channel that voids the plan on the first misprediction. Project page · Code · Paper · Integrity record The canonical release contains two… See the full description on the dataset page: https://huggingface.co/datasets/cveinnt/kepler-arc-agi-3-traces.tabularn<1K0 likes220 downloads21d agoHugging Face09sahil2801 /arc-agi-labelledtextn<1K2 likes89 downloads2y agoHugging Face10pxferna /ARC-AGI-v1textn<1K1 likes47 downloads1y agoHugging Face11Nabidnur /arc-agi-2-grids ARC-AGI-2 Grids — training + analysis corpus (NVARC-compatible) Companion dataset for the Kaggle ARC Prize 2026 (ARC-AGI-2) solver built on sorokin/qwen3_4b_grids15_sft139 + per-task rank-256 LoRA (NVARC lineage). Everything here is generated from public canonical data only (1,000 training / 120 evaluation tasks); no hidden competition data is included. Contents Path Rows Description train/train_tasks.jsonl 1,000 canonical training tasks (full I/O)… See the full description on the dataset page: https://huggingface.co/datasets/Nabidnur/arc-agi-2-grids.tabulartext-generation100K<n<1M0 likes41 downloads3d agoHugging Face12flaitenberger /arc_agi_1_augmentedtabular1M<n<10M0 likes34 downloads8mo agoHugging Face13iamjasonfeng /RPS-ARC-AGI-1-and-2This is the DPO dataset used in the RPS paper ( https://github.com/iamjasonfeng/RPS-Paper ) This dataset is based on the following dataset from Trelis: https://huggingface.co/datasets/Trelis/arc-agi-2-reasoning-5 text1K<n<10K0 likes32 downloads2mo agoHugging Face14tttx /r1-trajectories-arcagi-barctext1K<n<10K0 likes31 downloads2y agoHugging Face15omrisap /arc-agi-1-zloopvit-traces ARC-AGI-1 ZLoopViT Traces Canonical intermediate grid trajectories for all 400 training tasks in ARC-AGI-1. Each row corresponds to one original official train or test pair; generated augmentations are not included. Dataset contents 400 tasks 1,718 trajectories: 1,302 demonstration/train pairs and 416 test pairs 5,059 visible intermediate transitions Exact final-output validation on every official pair Important columns: task_id: official ARC task identifier… See the full description on the dataset page: https://huggingface.co/datasets/omrisap/arc-agi-1-zloopvit-traces.textimage-to-image1K<n<10K0 likes30 downloads1mo agoHugging Face16Nabidnur /arc-agi-2-cot-sft ARC-AGI-2 CoT-Solving SFT Dataset Companion to Nabidnur/arc-agi-2-grids. Grid-based chain-of-thought transcripts for the ARC-AGI-2 Kaggle competition, NVARC format compatible (Qwen chat template, {" "}-separated digit rows). Tiers program-verified — a transformation rule fitted ONLY on the demonstrations reproduces all of them exactly (verified=true). Hypothesis + per-pair verification + application. program-true (synthetic) — rules from generator ancestry… See the full description on the dataset page: https://huggingface.co/datasets/Nabidnur/arc-agi-2-cot-sft.text1K<n<10K0 likes28 downloads3d agoHugging Face17Nabidnur /arc-agi-2-synthetic-v1 ARC-AGI-2 Synthetic Curriculum v1 Program-generated ARC-like tasks with full ancestry records and independent verification (every task's rule must be re-fittable from its own demonstrations; 44 degenerate tasks rejected). Families (7): d4_transform, color_remap, tile_k, crop_bbox, symmetry_fill, hole_recolor, path_propagation. Counts family verified d4_transform 450 color_remap 450 tile_k 450 crop_bbox 450 symmetry_fill 439 hole_recolor 417… See the full description on the dataset page: https://huggingface.co/datasets/Nabidnur/arc-agi-2-synthetic-v1.text1K<n<10K0 likes24 downloads3d agoHugging Face18tttx /regular-arcagi-3k-r1-020925text1K<n<10K0 likes19 downloads2y agoHugging Face19omrisap /arc-agi-1-ruleloopvit-rules ARC-AGI-1 RuleLoopViT Rules This dataset contains one canonical, task-specific English rule for each of the 400 official ARC-AGI-1 training tasks. Rules were inferred only from official demonstration input/output pairs. Official test inputs, test outputs, and test traces were excluded from rule authoring. Each row includes: a concise standalone core_rule_text; a five-section full_rule_text; the corresponding structured sections; augmentation-aware references for colors… See the full description on the dataset page: https://huggingface.co/datasets/omrisap/arc-agi-1-ruleloopvit-rules.tabulartext-classificationn<1K0 likes16 downloads1mo agoHugging Face20pxferna /ARC-AGI-v1-5050textn<1K0 likes14 downloads1y agoHugging Face21flaitenberger /arc_agi_2_augmentedtabular1M<n<10M0 likes13 downloads8mo agoHugging Face22tttx /r1-masked-arcagi-v0text1K<n<10K0 likes5 downloads2y agoHugging Face23tttx /r1-masked-arcagi-v1text1K<n<10K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.