datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cybergymcybergymarvo-cybergym-2000
ARVO CyberGym-format 2000-task dataset
This dataset is shaped to be loaded by Harbor's CyberGym adapter.
It combines jm-rt/arvo-cybergym-1000 with the second 1000-task
small-target ARVO batch built outside the original CyberGym set.
cybergymarvo-cybergym-2000
ARVO CyberGym-format 2000-task dataset
This dataset is shaped to be loaded by Harbor's CyberGym adapter.
It combines jm-rt/arvo-cybergym-1000 with the second 1000-task
small-target ARVO batch built outside the original CyberGym set.
cybergymcybergymcybergym-tasks
CyberGym task ID splits
Task ID splits used in our work on CyberGym vulnerability-reproduction
benchmarks. Each row is a single task identifier (no inputs / no outputs);
this dataset is intended as a task-list manifest for downstream evaluation
scripts that fetch the actual task workspaces from the CyberGym distribution.
Configs
Config
Split
Rows
Description
full
train
1507
Every task in the CyberGym Level-1 release
train
train
300
Training pool used in our… See the full description on the dataset page: https://huggingface.co/datasets/Elfsong/cybergym-tasks.cybergymcybergymcybergym-lightcoredteam-cybergym-toy-taskscybergym-glm5
