datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cybergymarvo-cybergym-2000
ARVO CyberGym-format 2000-task dataset
This dataset is shaped to be loaded by Harbor's CyberGym adapter.
It combines jm-rt/arvo-cybergym-1000 with the second 1000-task
small-target ARVO batch built outside the original CyberGym set.
cybergymcybergym-e2earvo-cybergym-2000
ARVO CyberGym-format 2000-task dataset
This dataset is shaped to be loaded by Harbor's CyberGym adapter.
It combines jm-rt/arvo-cybergym-1000 with the second 1000-task
small-target ARVO batch built outside the original CyberGym set.
cybergym-servercybergymcybergym-server-binarycybergym-serverarvo-cybergym-2000
ARVO CyberGym-format 2000-task dataset
This dataset is shaped to be loaded by Harbor's CyberGym adapter.
It combines jm-rt/arvo-cybergym-1000 with the second 1000-task
small-target ARVO batch built outside the original CyberGym set.
cybergymcybergym-poccybergymcybergym-tasks
CyberGym task ID splits
Task ID splits used in our work on CyberGym vulnerability-reproduction
benchmarks. Each row is a single task identifier (no inputs / no outputs);
this dataset is intended as a task-list manifest for downstream evaluation
scripts that fetch the actual task workspaces from the CyberGym distribution.
Configs
Config
Split
Rows
Description
full
train
1507
Every task in the CyberGym Level-1 release
train
train
300
Training pool used in our… See the full description on the dataset page: https://huggingface.co/datasets/Elfsong/cybergym-tasks.cybergymcybergym-servercybergymcybergym-servercybergym-lighteasy-cybergym-datacybergym-server-binarycybergym-server-binarycoredteam-cybergym-toy-tasksCybergym-Experimentscoredteam-cybergym-toy-resultscybergym-glm5
