CoolFace
15 results

gemini3.1

AgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-tr87 ARC-AGI-3 tr87 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game tr87, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-tr87.tabularreinforcement-learningn<1K0 likes699 downloads1mo agoHugging FaceAgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-ar25 ARC-AGI-3 ar25 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game ar25, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-ar25.reinforcement-learning0 likes660 downloads1mo agoHugging FaceAgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-g50t ARC-AGI-3 g50t — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game g50t, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-g50t.tabularreinforcement-learningn<1K0 likes592 downloads1mo agoHugging FaceAgentNativeResearchLab /discoverphysics-gemini3.1-pro-ara DiscoverPhysics × Gemini 3.1 Pro High (Antigravity CLI) — 11-world ARA knowledge artifacts Agent-Native Research Artifacts (ARA) produced by a Gemini 3.1 Pro High (Antigravity CLI) coding-agent session solving all 11 worlds of the DiscoverPhysics scientific-discovery benchmark (seed 0, noise_frac 0.075, ≤16 experiment rounds), driven through the same harness-agnostic bridge and ARA scaffold as the sibling fable run. Official verdicts: 1/11 PASS — criteria and per-world numbers… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/discoverphysics-gemini3.1-pro-ara.0 likes454 downloads1mo agoHugging FaceAgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-s5i5 ARC-AGI-3 s5i5 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game s5i5, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-s5i5.reinforcement-learning0 likes427 downloads1mo agoHugging FaceAgentNativeResearchLab /arc-agi3-agy-gemini3.1pro-ls20 ARC-AGI-3 ls20 — Agent Trajectories (agy-gemini3.1pro) Gameplay trajectories from the harness×model pair agy-gemini3.1pro playing the ARC-AGI-3 game ls20, part of the ARA-as-world-model generalization experiment. The agent builds a structured world model (an Agent-Native Research Artifact) live during play and consults it to crack levels it cannot solve from cold exploration. One dataset repo per harness×model×game: sibling repos arc-agi3-<harness>-<model>-<game> hold the same… See the full description on the dataset page: https://huggingface.co/datasets/AgentNativeResearchLab/arc-agi3-agy-gemini3.1pro-ls20.reinforcement-learning0 likes405 downloads1mo agoHugging Face