arc-prize
arc_agi_v2_public_evalarc_agi_v1_public_evalARC Prize Foundation
Data associated with the output of models tested with https://github.com/arcprizeorg/model_baseline
arc_agi_2_human_testing
ARC-AGI-2 Human testing data
This file contains data from human testing sessions on ARC-AGI tasks.
Each row represents a single test attempt by a human participant on a specific task-test pair in the "Public Train" or "Public Eval" ARC-AGI-2 datasets. Not all tasks in the released "Public Train"
sets were tested, so these results are not comprehensive. This data does not include tasks from "Semi Private Evaluation" or "Private Evaluation"
Column Descriptions… See the full description on the dataset page: https://huggingface.co/datasets/arcprize/arc_agi_2_human_testing.arc_prize_public_eval
arc_prize_public_eval — evaluation data (OpenCompass format)
Bud Ecosystem eval mirror. OpenCompass-format eval data for arc_prize_public_eval (config arc_prize_public_eval_gen). Original source fchollet/ARC-AGI (GitHub); ARC-AGI-2 at arcprize/ARC-AGI-2 — license Apache-2.0, unchanged.
arc-prize-2024
