datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Agentic-Chain-of-Thought-Coding-SFT-Dataset
🤖 Agentic Coding CoT Dataset
A high-quality supervised fine-tuning (SFT) dataset for training agentic coding assistants with Chain-of-Thought reasoning capabilities.
📋 Dataset Description
This dataset was created by processing and distilling ~20GB of GitHub crawl data using Minimax-M2 to generate structured, reasoning-rich coding examples. Each sample demonstrates systematic problem-solving with explicit tool usage patterns.
🏗️ Assistant Data Structure… See the full description on the dataset page: https://huggingface.co/datasets/AlicanKiraz0/Agentic-Chain-of-Thought-Coding-SFT-Dataset.agentic-publication-protocol-dataset
APP compare-app benchmark
Paired reader conversations and blinded evaluations comparing an Agentic
Publication Protocol (APP) paper agent against a general repository-aware
agent, on 11 quantum-physics papers.
For each paper, a neutral reader asks the same scripted questions to both agents;
the two transcripts are anonymized and scored by a blinded evaluator on
accuracy, informativeness, grounding, and honesty (1-10).
Evaluator: Codex CLI, gpt-5.5, reasoning effort xhigh… See the full description on the dataset page: https://huggingface.co/datasets/phynics/agentic-publication-protocol-dataset.Agentic-SLS-Database
Agentic-SLS-Database
Canonical graph dataset of Inova Mk1 SLS printer entities: jobs, print sessions, print profiles, and objects (STL geometry). Each entity is its own HF config; relationships are encoded as ID references between rows.
Domain-specific datasets (e.g. ppak10/Agentic-SLS-ASTM) reference rows here by ID and may embed frozen snapshots of the referenced state.
Configs
Config
Description
Script
Output
jobs
One row per .s4a print job, with… See the full description on the dataset page: https://huggingface.co/datasets/ppak10/Agentic-SLS-Database.Agentic-Chain-of-Thought-Coding-SFT-Dataset-v1.1
🤖 Agentic Coding CoT Dataset v1.1
A high-quality supervised fine-tuning (SFT) dataset for training agentic coding assistants with Chain-of-Thought reasoning capabilities.
📋 Dataset Description
This dataset was created by processing and distilling ~20GB of GitHub crawl data using Minimax-M2 & MiniMax M2.1 to generate structured, reasoning-rich coding examples. Each sample demonstrates systematic problem-solving with explicit tool usage patterns.
🏗️ Assistant… See the full description on the dataset page: https://huggingface.co/datasets/AlicanKiraz0/Agentic-Chain-of-Thought-Coding-SFT-Dataset-v1.1.hendar-agentic-ai-dataset
Hendar Agentic AI Evaluation & Security Benchmark
A compact, expert-authored benchmark for evaluating trustworthy agentic AI systems across capability, tool use, retrieval, security, policy enforcement, multi-agent coordination and regression safety.
This dataset is a public companion to the Agentic AI Academy by Hendar Mawan, PhD. It is designed for evaluation, CI regression testing, red-team exercises and engineering education—not as a generic instruction-tuning corpus.… See the full description on the dataset page: https://huggingface.co/datasets/h0000w/hendar-agentic-ai-dataset.agentic-incident-response-20260905-dataset
Agentic Incident Response Orchestrator Synthetic Dataset
Summary
This dataset contains 14 training examples and 4
held-out examples for Production teams need agentic automation without allowing an LLM-style planner to execute unsafe remediation.
Every record is synthetic and includes:
input: query, event, or feature description
label: expected class, route, relation, or evidence category
context: synthetic supporting context
source: fictional source identifier… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/agentic-incident-response-20260905-dataset.xenon-mixed-agentic-datasetv1
Strategic Coding Traces MoE-Mix
This dataset is a highly curated mixture of coding agent traces designed specifically for Supervised Fine-Tuning (SFT) of Mixture-of-Experts (MoE) architecture models (e.g., Gemma-4-26B MoE).
Instead of using a single source or standard language distribution, this dataset employs Strategic Language Weighting. It routes the best coding agent data to specific language paradigms, allowing the MoE router to naturally learn to specialize: routing… See the full description on the dataset page: https://huggingface.co/datasets/el4/xenon-mixed-agentic-datasetv1.agentic-search-data
Agentic search — synthetic multi-hop retrieval dataset
JSONL artifacts for training and evaluating a retrieval agent across web, finance, legal, code, and science.
Files
File
Description
corpus.jsonl
Unique doc_id passages (supporting + distractor docs) with domain labels
sft_dataset.jsonl
Supervised fine-tuning tasks (~60% of tasks)
rl_dataset.jsonl
RL / GRPO-style prompts (~25%)
eval_dataset.jsonl
Held-out evaluation (~15%)
Splits are disjoint by… See the full description on the dataset page: https://huggingface.co/datasets/2796gauravc/agentic-search-data.APP1-Agentic-Safety-SFT-DataIKNN-Rl1-Dataset-Agentic-V2
IKNN-Rl1-Dataset-Agentic-V2
agentic-incident-response-20260826-dataset
Agentic Incident Response Orchestrator Synthetic Dataset
Summary
This dataset contains 14 training examples and 4
held-out examples for Production teams need agentic automation without allowing an LLM-style planner to execute unsafe remediation.
Every record is synthetic and includes:
input: query, event, or feature description
label: expected class, route, relation, or evidence category
context: synthetic supporting context
source: fictional source identifier… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/agentic-incident-response-20260826-dataset.Xenon-Mixed-Agentic-Dataset-v2
Xenon Mixed Agentic Dataset v2
This dataset is a highly curated mixture of coding agent traces designed specifically for Supervised Fine-Tuning (SFT) of Mixture-of-Experts (MoE) architecture models.
Instead of using a single source or standard language distribution, this dataset employs Strategic Language Weighting. It routes the best coding agent data to specific language paradigms, allowing an MoE router to naturally learn to specialize: routing Python/Bash tokens to Fable-5… See the full description on the dataset page: https://huggingface.co/datasets/el4/Xenon-Mixed-Agentic-Dataset-v2.cerebras_agentic_data1agentic-incident-response-20260717-dataset
Agentic Incident Response Orchestrator Synthetic Dataset
Summary
This dataset contains 14 training examples and 4
held-out examples for Production teams need agentic automation without allowing an LLM-style planner to execute unsafe remediation.
Every record is synthetic and includes:
input: query, event, or feature description
label: expected class, route, relation, or evidence category
context: synthetic supporting context
source: fictional source identifier… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/agentic-incident-response-20260717-dataset.agentic-incident-response-20260727-dataset
Agentic Incident Response Orchestrator Synthetic Dataset
Summary
This dataset contains 14 training examples and 4
held-out examples for Production teams need agentic automation without allowing an LLM-style planner to execute unsafe remediation.
Every record is synthetic and includes:
input: query, event, or feature description
label: expected class, route, relation, or evidence category
context: synthetic supporting context
source: fictional source identifier… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/agentic-incident-response-20260727-dataset.agentic-incident-response-20260816-dataset
Agentic Incident Response Orchestrator Synthetic Dataset
Summary
This dataset contains 14 training examples and 4
held-out examples for Production teams need agentic automation without allowing an LLM-style planner to execute unsafe remediation.
Every record is synthetic and includes:
input: query, event, or feature description
label: expected class, route, relation, or evidence category
context: synthetic supporting context
source: fictional source identifier… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/agentic-incident-response-20260816-dataset.Agentic-Chain-of-Thought-Coding-SFT-Dataset-v1.1
🤖 Agentic Coding CoT Dataset v1.1
A high-quality supervised fine-tuning (SFT) dataset for training agentic coding assistants with Chain-of-Thought reasoning capabilities.
📋 Dataset Description
This dataset was created by processing and distilling ~20GB of GitHub crawl data using Minimax-M2 & MiniMax M2.1 to generate structured, reasoning-rich coding examples. Each sample demonstrates systematic problem-solving with explicit tool usage patterns.
🏗️ Assistant… See the full description on the dataset page: https://huggingface.co/datasets/mepartha/Agentic-Chain-of-Thought-Coding-SFT-Dataset-v1.1.mirror-Agentic-Chain-of-Thought-Coding-SFT-Dataset
🤖 Agentic Coding CoT Dataset
A high-quality supervised fine-tuning (SFT) dataset for training agentic coding assistants with Chain-of-Thought reasoning capabilities.
📋 Dataset Description
This dataset was created by processing and distilling ~20GB of GitHub crawl data using Minimax-M2 to generate structured, reasoning-rich coding examples. Each sample demonstrates systematic problem-solving with explicit tool usage patterns.
🏗️ Assistant Data… See the full description on the dataset page: https://huggingface.co/datasets/alucent/mirror-Agentic-Chain-of-Thought-Coding-SFT-Dataset.mirror-Agentic-Chain-of-Thought-Coding-SFT-Dataset-v1.1
🤖 Agentic Coding CoT Dataset v1.1
A high-quality supervised fine-tuning (SFT) dataset for training agentic coding assistants with Chain-of-Thought reasoning capabilities.
📋 Dataset Description
This dataset was created by processing and distilling ~20GB of GitHub crawl data using Minimax-M2 & MiniMax M2.1 to generate structured, reasoning-rich coding examples. Each sample demonstrates systematic problem-solving with explicit tool usage patterns.
🏗️… See the full description on the dataset page: https://huggingface.co/datasets/alucent/mirror-Agentic-Chain-of-Thought-Coding-SFT-Dataset-v1.1.APP1-Agentic-Safety-SFT-Datapycoder-lfm2.5-agentic-datasetagentic-publication-protocol-dev-data
APP compare-app benchmark
Paired reader conversations and blinded evaluations comparing an Agentic
Publication Protocol (APP) paper agent against a general repository-aware
agent, on 11 public quantum-physics papers. This is the public-paper
subset reported in the APP paper's compare-app table.
For each paper, a neutral reader asks the same scripted questions to both agents;
the two transcripts are anonymized and scored by a blinded evaluator on
accuracy, informativeness… See the full description on the dataset page: https://huggingface.co/datasets/LionSR/agentic-publication-protocol-dev-data.xDAN-Agentic-Data-Instruct-v1-sampleagentic_conversationsagentic_datahow-agentic-m0-dataagentic-planner-datasettest-dataset-agentic
