datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Sequential-Math
RecursiveMAS Sequential-Math
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Sequential-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Sequential-Math
Original file
Sequential-Math.json
Collaboration style
Sequential-Style
Used for
sequential math inner agents and outer RecursiveLink… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Sequential-Math.recursive-task-synthesis-glm-5.3-rollouts
GLM 5.3 agentic rollouts on Recursive-Task-Synthesis
This dataset catalogs the full collection made from the pinned
Recursive-Task-Synthesis dataset revision
be44f96808d5a9b599d5cb024341ff00091adeb7. The repository includes approximately 260.5 GiB of trajectory payload tar shards.
Contents at a glance
Item
Count
Source tasks considered
37,284
Source candidates inspected
19,368
Converted tasks after source filters
18,600
Tasks passing gold… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/recursive-task-synthesis-glm-5.3-rollouts.pubmedqa-recursive-llm-degradation-qwen2.5-0.5b
PubMedQA Recursive LLM Degradation — Qwen2.5-3B
This repository contains synthetic biomedical question-answering data
and model predictions generated as part of a study of recursive
fine-tuning and model degradation.
Base Model
Qwen/Qwen2.5-3B
Source Dataset
The experiments use the PubMedQA dataset:
qiaoxin/PubMedQA
This repository contains generated/derived research artifacts and does
not redistribute the original PubMedQA dataset in its entirety.… See the full description on the dataset page: https://huggingface.co/datasets/chrislimbe/pubmedqa-recursive-llm-degradation-qwen2.5-0.5b.pubmedqa-recursive-llm-degradation-qwen2.5-3b
PubMedQA Recursive LLM Degradation — Qwen2.5-3B
This repository contains synthetic biomedical question-answering data
and model predictions generated as part of a study of recursive
fine-tuning and model degradation.
Base Model
Qwen/Qwen2.5-3B
Source Dataset
The experiments use the PubMedQA dataset:
qiaoxin/PubMedQA
This repository contains generated/derived research artifacts and does
not redistribute the original PubMedQA dataset in its entirety.… See the full description on the dataset page: https://huggingface.co/datasets/chrislimbe/pubmedqa-recursive-llm-degradation-qwen2.5-3b.recursive-lines
Recursive Lines: A Dual-Track Adversarial Benchmark
Recursive Lines is a diagnostic suite for detecting "High-Agency Deception" in Large Language Models. It serves as the reference implementation for the Constraint Cascade Model (FAccT 2026) and the Agency Index metric.
1. Overview
Current LLM benchmarks measure capability (MMLU) or safety (Refusal). They fail to measure Agency—the thermodynamic distinction between stochastic error (hallucination) and strategic intent… See the full description on the dataset page: https://huggingface.co/datasets/OstensibleParadox/recursive-lines.openclaw-recursive-study-data
OpenClaw Recursive Repository Study Data
Synthetic repository-study data generated against
openclaw/openclaw at commit
da228660306b55a9cce3b973946f3aacfc515848. The source repository is MIT licensed.
This release contains exploration questions, tool-using study trajectories,
recursive notes, full recall-rewritten trajectories, and recall-to-action
training examples. Nested chat/tool objects are stored as JSON strings to keep
the schema stable and can be decoded with json.loads.… See the full description on the dataset page: https://huggingface.co/datasets/aviralku/openclaw-recursive-study-data.Distillation-Code
RecursiveMAS Distillation-Code
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Distillation-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Distillation-Code
Original file
Distillation-Code.json
Collaboration style
Distillation-Style
Used for
expert/learner code inner agents and outer… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Distillation-Code.Sequential-Code
RecursiveMAS Sequential-Code
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Sequential-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Sequential-Code
Original file
Sequential-Code.json
Collaboration style
Sequential-Style
Used for
sequential code inner agents and outer RecursiveLink… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Sequential-Code.recursive-memory-perfectblend-coding
Frozen PerfectBlend + Coding training data
Private migration snapshot of the raw, Qwen-generated training trajectories used
by the all-turn residual Compressor dataset frozen on 2026-09-06. This is not
the untouched upstream PerfectBlend dataset or a newly generated corpus.
Corpus
Trajectories
Generated assistant responses
Raw bytes
PerfectBlend / xhigh
40,596
51,039
549,854,699
Coding
37,560
86,880
1,274,359,351
Total
78,156
137,919
1,824,214,050
The mixture… See the full description on the dataset page: https://huggingface.co/datasets/mocoV3/recursive-memory-perfectblend-coding.Mixture-Math
RecursiveMAS Mixture-Math
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Mixture-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Mixture-Math
Original file
Mixture-Math.json
Collaboration style
Mixture-Style
Used for
math specialist inner agent training
Split
train
Rows
1904… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Mixture-Math.Vinayak-Multistep-Recursive-Reasoning-Benchmark
Vinayak Multistep Recursive Reasoning Benchmark (VMRRB)
Overview
The Vinayak Multistep Recursive Reasoning Benchmark (VMRRB) is a large-scale prompt-based benchmark designed to evaluate advanced reasoning, recursive dependency resolution, encrypted task traversal, and robustness capabilities of frontier AI systems.
The benchmark evaluates a model's ability to:
Perform recursive multistep reasoning
Resolve interdependent question chains
Execute encrypted dependency… See the full description on the dataset page: https://huggingface.co/datasets/bepipeV/Vinayak-Multistep-Recursive-Reasoning-Benchmark.Distillation-Math
RecursiveMAS Distillation-Math
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Distillation-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Distillation-Math
Original file
Distillation-Math.json
Collaboration style
Distillation-Style
Used for
expert/learner math inner agents and outer… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Distillation-Math.Deliberation
RecursiveMAS Deliberation
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Deliberation-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Deliberation
Original file
Deliberation.json
Collaboration style
Deliberation-Style
Used for
reflector/tool-caller inner agents and outer RecursiveLink… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Deliberation.Mixture-Science
RecursiveMAS Mixture-Science
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Mixture-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Mixture-Science
Original file
Mixture-Science.json
Collaboration style
Mixture-Style
Used for
science specialist inner agent training
Split
train
Rows… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Mixture-Science.nanochat-connectk-recursive-frames-v1
Nanochat Connect-K Recursive frames v1
This is a static, deterministic pretraining rendering of the validated
recursive Connect-K disjoint V3 proofs. Recursive rows are runtime-aligned
frame segments: a new document begins whenever strict execution enters a child
frame or resumes a parent after replacing its call block with the copied child
return. The paired recursive and flat-CoT repositories use the exact same
ordered roots and canonical source proofs.
Their token streams and… See the full description on the dataset page: https://huggingface.co/datasets/SolidSnake123/nanochat-connectk-recursive-frames-v1.Vinayak-Multistep-Recursive-Reasoning-Benchmark
Vinayak Multistep Recursive Reasoning Benchmark (VMRRB)
Overview
The Vinayak Multistep Recursive Reasoning Benchmark (VMRRB) is a large-scale prompt-based benchmark designed to evaluate advanced reasoning, recursive dependency resolution, encrypted task traversal, and robustness capabilities of frontier AI systems.
The benchmark evaluates a model's ability to:
Perform recursive multistep reasoning
Resolve interdependent question chains
Execute encrypted… See the full description on the dataset page: https://huggingface.co/datasets/SavantCapital/Vinayak-Multistep-Recursive-Reasoning-Benchmark.Mixture-Outer
RecursiveMAS Mixture-Outer
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Mixture-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Mixture-Outer
Original file
Mixture-Outer.json
Collaboration style
Mixture-Style
Used for
mixture outer RecursiveLink training
Split
train
Rows
4904… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Mixture-Outer.Mixture-Code
RecursiveMAS Mixture-Code
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Mixture-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Mixture-Code
Original file
Mixture-Code.json
Collaboration style
Mixture-Style
Used for
code specialist inner agent training
Split
train
Rows
2000… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Mixture-Code.Mixture-Summarizer
RecursiveMAS Mixture-Summarizer
Project Page | Code | Paper
We introduce RecursiveMAS, a multi-agent framework that scales agent collaboration through latent-space recursion. This dataset contains training examples for the Mixture-Style setting.
Dataset Details
Item
Description
Dataset
RecursiveMAS/Mixture-Summarizer
Original file
Mixture-Summarizer.json
Collaboration style
Mixture-Style
Used for
summarizer inner agent training
Split
train… See the full description on the dataset page: https://huggingface.co/datasets/RecursiveMAS/Mixture-Summarizer.
