datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
public-agent-coordination-artifacts
Public Agent Coordination Artifacts
Real, complete edits and posts that AI agents left on public wikis and paste sites —
collected as open evidence for studying how autonomous agents use shared online spaces to
remember things, signal each other, and coordinate. It's the behavior spotlighted by the
mid-2026 OpenAI–Hugging Face agent incident,
here as raw public data researchers can actually inspect — plus a small, hand-reviewed map
of how specific artifacts relate.… See the full description on the dataset page: https://huggingface.co/datasets/leonidas1712/public-agent-coordination-artifacts.agent-code-rl-artifacts
Agent Code RL Artifacts
Recovered process data from a code-generation Agent project covering SFT,
Monte Carlo rollout, process reward modeling, and veRL GRPO. This repository
contains benchmark-derived records and AI-generated content; it is not a
human-authored-only dataset.
Related SFT adapter:
keryszhan/qwen2.5-coder-7b-code-plan-sft.
Data stages
Config
Purpose
Important boundary
splits
Canonical HumanEval/MBPP-derived task splits
grpo_evaluation is… See the full description on the dataset page: https://huggingface.co/datasets/keryszhan/agent-code-rl-artifacts.
