CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AgentNativeResearchLab /arc-agi3-codex-gpt5.6sol-ls20 ARC-AGI-3 ls20 — Agent Trajectories (codex-gpt5.6sol) Gameplay trajectories from the harness×model pair codex-gpt5.6sol playing the ARC-AGI-3 game ls20, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-codex-gpt5.6sol-ls20.textreinforcement-learning0 likes1.5k downloads25d agoHugging Face02schema-harness /arc-agi-3-schema-traces ARC-AGI-3 Schema Gameplay Trajectories This release contains 50 ARC-AGI-3 gameplay trajectories and a dependency-free scoring utility. The trajectories are split evenly across two collections: gpt_5_6_sol/: 25 GPT-5.6 Sol trajectories. claude_fable_opus/: 25 trajectories from Claude Opus 4.8 and Claude Fable 5. Each trajectory directory includes run.json, a streamed events.jsonl event log, sanitized session data, snapshots, and the shareable text/image files produced during… See the full description on the dataset page: https://huggingface.co/datasets/schema-harness/arc-agi-3-schema-traces.tabularn<1K38 likes1.4k downloads2mo agoHugging Face03AgentNativeResearchLab /arc-agi3-kimi-k2.7-ar25 ARC-AGI-3 ar25 — Agent Trajectories (kimi-k2.7) Gameplay trajectories from the harness×model pair kimi-k2.7 playing the ARC-AGI-3 game ar25, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game played by… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-kimi-k2.7-ar25.tabularreinforcement-learningn<1K0 likes1.1k downloads25d agoHugging Face04AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-tr87 ARC-AGI-3 tr87 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game tr87, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-tr87.tabularreinforcement-learningn<1K0 likes699 downloads1mo agoHugging Face05AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-g50t ARC-AGI-3 g50t — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game g50t, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-g50t.tabularreinforcement-learningn<1K0 likes592 downloads1mo agoHugging Face06magic-sword /arc_agi_3_public_demo_human_testing Dataset Card for ARC-AGI 3 Public Demo Human Testing Dataset Summary This dataset contains human gameplay logs and trajectories from the ARC-AGI 3 public demo. It is a fully open-source dataset created by the ARC Prize. The primary purpose of publishing this dataset on Hugging Face is to make it easily accessible and convenient for participants in the Kaggle ARC Prize 2026 Competition. The implementation and source code used to process and upload this dataset to… See the full description on the dataset page: https://huggingface.co/datasets/magic-sword/arc_agi_3_public_demo_human_testing.tabularreinforcement-learningn<1K1 likes520 downloads4mo agoHugging Face07JBrightmanAI /arc-agi-3-schema-traces ARC-AGI-3 Schema Gameplay Trajectories This release contains 50 ARC-AGI-3 gameplay trajectories and a dependency-free scoring utility. The trajectories are split evenly across two collections: gpt_5_6_sol/: 25 GPT-5.6 Sol trajectories. claude_fable_opus/: 25 trajectories from Claude Opus 4.8 and Claude Fable 5. Each trajectory directory includes run.json, a streamed events.jsonl event log, sanitized session data, snapshots, and the shareable text/image files produced during… See the full description on the dataset page: https://huggingface.co/datasets/JBrightmanAI/arc-agi-3-schema-traces.tabularn<1K0 likes420 downloads2mo agoHugging Face08AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-su15 ARC-AGI-3 su15 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game su15, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-su15.tabularreinforcement-learningn<1K0 likes383 downloads1mo agoHugging Face09ritwika96 /arc-agi-3-full-curriculum ARC-AGI-3 full transformation curriculum This release contains 444,000 deterministic multimodal episodes across five curriculum configs: stateful: 168,000 episodes over 12 stateful causal families; search: 76,000 episodes over 6 graph, search, and objective families; physics: 96,000 episodes over 8 object-physics families. composition: 64,000 episodes over 8 pairwise mechanic templates, with complete validation and test pairings absent from training. ambiguity: 40,000 episodes… See the full description on the dataset page: https://huggingface.co/datasets/ritwika96/arc-agi-3-full-curriculum.imageimage-text-to-text100K<n<1M1 likes324 downloads2d agoHugging Face10AgentNativeResearchLab /arc-agi3-cc-fable5-ft09 ARC-AGI-3 ft09 — Agent Trajectories (cc-fable5) Gameplay trajectories from the harness×model pair cc-fable5 playing the ARC-AGI-3 game ft09, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game played by… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-cc-fable5-ft09.textreinforcement-learning1K<n<10K0 likes317 downloads25d agoHugging Face11fredericowieser /arc-agi-3-wm-traces ARC-AGI-3 World Model Traces This dataset contains ARC-AGI-3 transition traces in the same parquet schema used by HHazard/arc-agi-3. Each row is one environment transition: state, game_id, level_id, action_id, action_args, next_state, level_done, frame_idx, origin, transformation, player state and next_state are 64x64 ARC grids stored as nested integer arrays. action_id is the ARC-AGI-3 action kind; click actions use action_args.x and action_args.y. Splits… See the full description on the dataset page: https://huggingface.co/datasets/fredericowieser/arc-agi-3-wm-traces.tabularreinforcement-learning10M<n<100M0 likes311 downloads3mo agoHugging Face12cveinnt /kepler-arc-agi-3-traces Kepler 1.0 ARC-AGI-3 trace corpus Run artifacts from Kepler 1.0, an open-source agent harness for the 25 public ARC-AGI-3 games. A stock CLI coding agent encodes its theory of each game as an executable world_model.py, certifies it against the full recorded interaction history, plans inside the certified model, and acts through a guarded channel that voids the plan on the first misprediction. Project page · Code · Paper · Integrity record The canonical release contains two… See the full description on the dataset page: https://huggingface.co/datasets/cveinnt/kepler-arc-agi-3-traces.tabularn<1K0 likes216 downloads20d agoHugging Face13ritwika96 /arc-agi-3-transformation-induction ARC-AGI-3 transformation induction — scene-v2 phase 1 This release contains 130,000 deterministic multimodal episodes: 104,000 train 12,000 validation 14,000 test 3,250 latent mechanic identities 40 visual skins per identity The old seed-v1 release was a publication-pipeline fixture and is not training data. Scene-v2 uses persistent 64x64 game boards with arenas, panels, palettes, lattices, targets, HUD-like regions and coherent distractors. Why compact specs… See the full description on the dataset page: https://huggingface.co/datasets/ritwika96/arc-agi-3-transformation-induction.imageimage-text-to-text100K<n<1M0 likes161 downloads2d agoHugging Face14zarczynski /arc-agi-3-publictextn<1K0 likes122 downloads6mo agoHugging Face15HHazard /arc-agi-3tabular10K<n<100K1 likes114 downloads4mo agoHugging Face16AgentNativeResearchLab /arc-agi3-phase2-ar25 ARC-AGI-3 ar25 — ARA snapshots truncated by absolute inference-cost budget Phase-2 dataset. 6 agents (harness×model) each played ar25, continuously crystallizing a structured world model (Agent-Native Research Artifact). Here each agent's ARA is sliced at fixed absolute cumulative inference-cost budgets — $13/$27/$51/$76 — so you can ask "for a spend of $B, what world model has each agent built?" Structure cost_05usd/ cost_15usd/ cost_40usd/ cost_75usd/ <agent>/… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-phase2-ar25.tabularothern<1K0 likes114 downloads1mo agoHugging Face17AgentNativeResearchLab /arc-agi3-phase2-g50t ARC-AGI-3 g50t — ARA snapshots truncated by absolute inference-cost budget Phase-2 dataset. 6 agents (harness×model) each played g50t, continuously crystallizing a structured world model (ARA). Each agent's ARA is sliced at fixed absolute cumulative inference-cost budgets $13/$27/$51/$76 (data-driven natural cost-cluster boundaries). Structure cost_13usd/ cost_27usd/ cost_51usd/ cost_76usd/ <agent>/ ara/ frontier.md meta.json Price table (per… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-phase2-g50t.textother1K<n<10K0 likes109 downloads1mo agoHugging Face18dnhkng /arc-agi-3-teaching-suite ARC-AGI-3 teaching suite 100 grid-reasoning tasks across 16 puzzle families, with two splits: split rows answers in this dataset why taught 80 yes 16 families x 5 worked examples, each documented in INSTRUCTIONS.md and reasoned through in CHAIN_OF_THOUGHT.md heldout 20 no 16 unseen instances of the same families, plus 4 tasks that are provably not solvable by reasoning from their examples The teaching material is what makes this a teaching suite rather than a… See the full description on the dataset page: https://huggingface.co/datasets/dnhkng/arc-agi-3-teaching-suite.tabularothern<1K2 likes108 downloads4d agoHugging Face19HHazard /big-arc-agi-3tabular10M<n<100M0 likes81 downloads3mo agoHugging Face20guanning /arc-agi-3-schema-traces-gpt56gated ARC-AGI-3 Schema Gameplay Trajectories — GPT-5.6 Sol This release contains every gpt-5.6-sol gameplay trajectory produced on our cluster with the world_model_v5 agent harness — 100 runs across the 25 public ARC-AGI-3 games — plus a dependency-free scoring utility. It is the GPT-5.6 Sol member of a family built by the same harness and the same sanitizer, so trajectories can be compared game by game: arc-agi-3-schema-traces-fable5 — Claude Fable 5, best per game (25)… See the full description on the dataset page: https://huggingface.co/datasets/guanning/arc-agi-3-schema-traces-gpt56.tabularreinforcement-learningn<1K0 likes81 downloads5d agoHugging Face21guanning /arc-agi-3-schema-traces-opus48gated ARC-AGI-3 Schema Gameplay Trajectories — Claude Opus 4.8 This release contains the best claude-opus-4-8 / max trajectory for each of the 25 public ARC-AGI-3 games, plus a dependency-free scoring utility. It is the Opus 4.8 counterpart of arc-agi-3-schema-traces-fable5, produced by the same agent harness (world_model_v5) and the same sanitizer, so the two collections can be compared game by game. Each trajectory directory includes run.json, a streamed events.jsonl event log… See the full description on the dataset page: https://huggingface.co/datasets/guanning/arc-agi-3-schema-traces-opus48.tabularreinforcement-learningn<1K0 likes47 downloads5d agoHugging Face22Asap7772 /arc-agi-mixed-max4096-qwenabsgen-all-flat-train-Qwen3-4B-Thinking-2507-samp1-abs-alltext100K<n<1M0 likes37 downloads1y agoHugging Face23guanning /arc-agi-3-schema-traces-gpt56-xhighgated ARC-AGI-3 Schema Gameplay Trajectories — GPT-5.6 Sol (xhigh) The best gpt-5.6-sol trajectory at xhigh reasoning effort for each of the 25 public ARC-AGI-3 games, produced with the world_model_v5 agent harness. This release exists to make the cross-model comparison single-effort on all sides. Its siblings are each one model at one effort, but the gpt-5.6-sol collection in arc-agi-3-schema-gameplay is a mix of xhigh and max (16 games + 9 games), so it is not directly comparable to… See the full description on the dataset page: https://huggingface.co/datasets/guanning/arc-agi-3-schema-traces-gpt56-xhigh.tabularreinforcement-learningn<1K0 likes35 downloads3d agoHugging Face24Asap7772 /arc-agi-mixed-max4096-qwenabsgen-all-flat-train-Qwen3-4B-Thinking-2507-samp1-abs-all-validtabular100K<n<1M0 likes30 downloads1y agoHugging Face25AnsteyGeorge /arc-agi-3-publictextn<1K0 likes16 downloads4mo agoHugging Face26Asap7772 /arc-agi-mixed-max4096-qwenabsgen-all-flat-train-Qwen3-4B-Thinking-2507-samp1-abs-4of16text1K<n<10K0 likes14 downloads1y agoHugging Face27Asap7772 /arc-agi-mixed-barc-train-Qwen3-4B-samp16-alltext1K<n<10K0 likes13 downloads1y agoHugging Face28Asap7772 /arc-agi-all-Qwen3-4B-samp16-4of8textn<1K0 likes12 downloads1y agoHugging Face29Asap7772 /arc-agi-mixed-barc-train-Qwen3-4B-samp16-all-validtabular1K<n<10K0 likes11 downloads1y agoHugging Face30Asap7772 /arc-agi-mixed-max4096-qwenabsgen-all-flat-train-Qwen3-4B-Thinking-2507-samp1-abs-10of16text1K<n<10K0 likes11 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.