CoolFace
18 results

arc-agi-3

AgentNativeResearchLab /arc-agi3-codex-gpt5.5-su15 ARC-AGI-3 su15 — Agent Trajectories (codex-gpt5.5) Gameplay trajectories from the harness×model pair codex-gpt5.5 playing the ARC-AGI-3 game su15, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-codex-gpt5.5-su15.reinforcement-learning0 likes4.1k downloads24d agoHugging FaceAgentNativeResearchLab /arc-agi3-kimi-k2.7-su15 ARC-AGI-3 su15 — Agent Trajectories (kimi-k2.7) Gameplay trajectories from the harness×model pair kimi-k2.7 playing the ARC-AGI-3 game su15, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game played by… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-kimi-k2.7-su15.reinforcement-learning0 likes3.4k downloads24d agoHugging FaceAgentNativeResearchLab /arc-agi3-codex-gpt5.5-s5i5 ARC-AGI-3 s5i5 — Agent Trajectories (codex-gpt5.5) Gameplay trajectories from the harness×model pair codex-gpt5.5 playing the ARC-AGI-3 game s5i5, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-codex-gpt5.5-s5i5.reinforcement-learning0 likes2.6k downloads24d agoHugging FaceAgentNativeResearchLab /arc-agi3-codex-gpt5.5-r11l ARC-AGI-3 r11l — Agent Trajectories (codex-gpt5.5) Gameplay trajectories from the harness×model pair codex-gpt5.5 playing the ARC-AGI-3 game r11l, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-codex-gpt5.5-r11l.reinforcement-learning0 likes1.7k downloads24d agoHugging FaceAgentNativeResearchLab /arc-agi3-kimi-k2.7-ls20 ARC-AGI-3 ls20 — Agent Trajectories (kimi-k2.7) Gameplay trajectories from the harness×model pair kimi-k2.7 playing the ARC-AGI-3 game ls20, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game played by… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-kimi-k2.7-ls20.reinforcement-learning0 likes1.6k downloads24d agoHugging FaceAgentNativeResearchLab /arc-agi3-kimi-k2.7-g50t ARC-AGI-3 g50t — Agent Trajectories (kimi-k2.7) Gameplay trajectories from the harness×model pair kimi-k2.7 playing the ARC-AGI-3 game g50t, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same game played by… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-kimi-k2.7-g50t.reinforcement-learning2 likes1.6k downloads24d agoHugging Face