datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qwen3-8b-codi-multihop-recall-data
CODI training data — multi-hop recall & pointer-chase (single-token-node reasoning)
The training data + generators + load-bearing eval code for two Qwen3-8B CODI latent-reasoning organisms:
cds-jb/qwen3-8b-codi-multihop-recall and
cds-jb/qwen3-8b-codi-pointer-chase.
Both tasks are single-token-node serial-reasoning problems: every intermediate and the final answer is a
single token (in both the Qwen3 and Gemma3 tokenizers), so each CODI latent can in principle be read with a… See the full description on the dataset page: https://huggingface.co/datasets/cds-jb/qwen3-8b-codi-multihop-recall-data.agentsim-atc-multihop
AgentSim Agent-Trace Corpus — Multi-hop
A multi-hop sibling of the AgentSim Agent-Trace Corpus
(agentsim-atc)
with an evolved schema designed for student model distillation.
1 490 accepted SFT trajectories plus 2 980 step-level DPO preference
pairs, generated over 5 multi-hop QA datasets through a 7-action agentic
executor with an Always-Search Policy filter.
This corpus accompanies a follow-up technical report to "AgentSim: A
Platform for Verifiable Agent-Trace Simulation"… See the full description on the dataset page: https://huggingface.co/datasets/searchsim/agentsim-atc-multihop.
