datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
devopsbench-100
DevOpsBench-100
DevOpsBench-100 is a synthetic long-horizon software-engineering / SRE agent
benchmark: 100 tasks over one executable world ("NovaCart", a mid-size
e-commerce SaaS) with 72 SQLite tables,
1451 seeded rows, a 38-file monorepo with 417 commits,
and 97 MCP tools spanning a first-party engineering stack
(tickets, PRs, CI, deployments, canaries, migrations, feature flags, metrics,
alerts, incidents, chat, knowledge base) plus deliberately disagreeing
vendor-shaped… See the full description on the dataset page: https://huggingface.co/datasets/SamuelChien821/devopsbench-100.stack-v3-devops
The Stack v3 DevOps Corpus
13,234,862 complete infrastructure units extracted from
The Stack v3,
grouped into seven classes and gated on content rather than popularity.
A unit is not a file, it is the thing an engineer would actually run: a Helm chart
arrives with its Chart.yaml, values.yaml and every template; a Terraform module
with all of its .tf files; an Ansible role with its tasks, defaults and handlers.
That is only possible because The Stack v3 groups rows by repository… See the full description on the dataset page: https://huggingface.co/datasets/Helmcode/stack-v3-devops.devops-incident-response
Dataset Card for DevOps Incident Response Dataset
Dataset Description
Dataset Summary
The DevOps Incident Response Dataset is a comprehensive collection of real-world-style DevOps incidents, troubleshooting scenarios, and resolution procedures. This dataset is designed to help train AI models for DevOps assistance, incident response automation, and technical troubleshooting education.
Each incident includes:
Detailed incident description and symptoms… See the full description on the dataset page: https://huggingface.co/datasets/Snaseem2026/devops-incident-response.omnimcp_devops_cloud_teaser
🔬 INSPECT THE DEEPSEEK-R1 REASONING CHAIN LIVE:
Zero hallucinations. Null syntax errors. 100% AST compiler validated.🌐 Live Interactive Reasoning & Code Inspector: https://emgena.com/trainingslager🎁 Claim your Free Starter Kit (Code: STARTER100): https://emgena.com/trainingslager🏷️ Launch Discount: Get 20 € OFF any 500-incident production suite with code LAUNCH20!
📜 Enterprise Compliance: EU AI Act Articles 50 & 53 certified • 100% DSGVO / GDPR clean • Commercial EULA… See the full description on the dataset page: https://huggingface.co/datasets/emgena/omnimcp_devops_cloud_teaser.devops-qa-dataset
DevOps Q&A Dataset v1.0
Overview
High-quality dataset of 25,670 DevOps technical examples collected from GitHub repositories, Stack Exchange, and official documentation.
Statistics
Total examples: 25,670
Average quality score: ~0.82
Unique (deduplicated): High accuracy via MD5
Categories: Docker, Kubernetes, CI/CD, Cloud, Linux, Terraform, Ansible
Sources: StackExchange/HuggingFace (70%), GitHub Repositories (29%), Official Documentation (~1%)
Use… See the full description on the dataset page: https://huggingface.co/datasets/Skilln/devops-qa-dataset.devops-kubernetes-sft-100k
DevOps and Kubernetes SFT 100K
A synthetic supervised fine-tuning dataset of 100,000 high-quality DevOps and Kubernetes conversations designed to train AI assistants capable of supporting platform engineers, SREs, and DevOps practitioners.
Dataset Description
This dataset covers production-grade Kubernetes operations, cloud infrastructure, CI/CD pipelines, GitOps workflows, and platform engineering across 13 specialized categories. Each record follows the ShareGPT… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/devops-kubernetes-sft-100k.devops-cloud-instruction-dataset
DevOps & Cloud Infrastructure Dataset
Professional instruction-response pairs for DevOps engineers covering Kubernetes, Docker, Terraform, CI/CD, and cloud services (AWS, Azure).
Dataset Details
Dataset Description
This is a high-quality instruction-tuning dataset focused on Devops Cloud topics. Each entry includes:
A clear instruction/question
Optional input context
A detailed response/solution
Chain-of-thought reasoning process
Curated by: CloudKernel.IO… See the full description on the dataset page: https://huggingface.co/datasets/bernabepuente/devops-cloud-instruction-dataset.CodeFuse-DevOps-EvalDevOps-Eval is a comprehensive chinese evaluation suite specifically designed for foundation models in the DevOps field. It consists of 5977 multi-choice questions spanning 55 diverse categories. Please visit our website and GitHub for more details.
Each category consists of two splits: dev, and test. The dev set per subject consists of five exemplars with explanations for few-shot evaluation. And the test set is for model evaluation. Labels on the test split are released, users can evaluate… See the full description on the dataset page: https://huggingface.co/datasets/codefuse-ai/CodeFuse-DevOps-Eval.devops-sft-dataset
DevOps SFT Instruction Dataset
This dataset contains 8,076 high-quality instruction-response pairs specifically generated for fine-tuning a DevOps domain-specialized language model. It was used in the Supervised Fine-Tuning (SFT) phase of the Ulysses model training pipeline.
Dataset Description
Instructions were generated using the Gemini API (gemini-2.0-flash) and Ollama (qwen2.5-coder:7b) by feeding chunks of official DevOps documentation and GitHub repositories… See the full description on the dataset page: https://huggingface.co/datasets/jalpan04/devops-sft-dataset.devops-kubernetes-iac-sft-dpo-2026
⚙️ Enterprise DevOps AI, Kubernetes SRE & IaC SFT/DPO Dataset (2026)
High-precision multi-turn instruction tuning and preference optimization dataset with step-by-step SRE root-cause Chain-of-Thought (<thought>) diagnostic trees for fine-tuning LLMs (Llama-3.3, Qwen-2.5-Coder, DeepSeek-R1-Distill, Mistral) into Senior Site Reliability Engineers (SRE), Principal Cloud Architects, and DevSecOps Specialists.
📊 Dataset Architecture & Highlights
Multi-Turn SRE… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/devops-kubernetes-iac-sft-dpo-2026.Devops_LLMdevops-predictive-logs
🔮 DevOps Predictive Logs Dataset
A synthetic dataset of realistic DevOps log sequences for training and benchmarking predictive failure models.
📊 Dataset Summary
This dataset contains realistic DevOps log scenarios covering common infrastructure failure patterns. Each log entry includes metadata about the failure scenario, severity, and time-to-failure, making it ideal for training predictive models.
Total Logs: ~150+ entriesScenarios: 10 unique failure patternsFormat:… See the full description on the dataset page: https://huggingface.co/datasets/Snaseem2026/devops-predictive-logs.deepfabric-devops-reasoning-traces
Dataset Description
DevOps reasoning traces
Dataset Details
Created by: Always Further
License: CC BY 4.0
Language(s): [English
Dataset Size: 10050
Data Splits
[train]
Dataset Creation
This dataset was created using DeepFabric, an open-source tool for generating high-quality training datasets for AI models.
Usage
from datasets import load_dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/nolabs/deepfabric-devops-reasoning-traces.SO-Python_QA-System_Administration_and_DevOps_classstackexchange_devopsbio-devops-synthetic-instructions
Bio-DevOps Synthetic Instructions
This dataset contains synthetic instruction-following examples for biomedical-style data-engineering and scientific-computing workflows.
It was created for educational and portfolio use as part of a LoRA/QLoRA fine-tuning project using Qwen/Qwen2.5-Coder-7B-Instruct.
Related model:
AiLLMBS/qwen25-coder-bio-devops-lora
Dataset Contents
The dataset includes synthetic examples for:
Python CSV validation
pandas duplicate checks
bash… See the full description on the dataset page: https://huggingface.co/datasets/AiLLMBS/bio-devops-synthetic-instructions.smoltrace-devops-tasks
SMOLTRACE Synthetic Dataset
This dataset was generated using the TraceMind MCP Server's synthetic data generation tools.
Dataset Info
Tasks: 100
Format: SMOLTRACE evaluation format
Generated: AI-powered synthetic task generation
Usage with SMOLTRACE
from datasets import load_dataset
# Load dataset
dataset = load_dataset("MCP-1st-Birthday/smoltrace-devops-tasks")
# Use with SMOLTRACE
# smoltrace-eval --model openai/gpt-4 --dataset-name… See the full description on the dataset page: https://huggingface.co/datasets/MCP-1st-Birthday/smoltrace-devops-tasks.devops-dataset-v1devops_predictive_logs
DevOps Predictive Logs Dataset (TsFile)
Apache TsFile version of Snaseem2026/devops-predictive-logs.
Overview
A synthetic dataset of realistic DevOps log sequences for training and
benchmarking predictive failure models. Each log entry describes one
infrastructure observation (service, pod, level, message) inside one of 10
failure scenarios together with its incident metadata: severity, whether the
pod eventually fails, and the time-to-failure in minutes.
Rows:… See the full description on the dataset page: https://huggingface.co/datasets/THULab/devops_predictive_logs.devops-promql-sre-curated-600
🚀 DevOps SRE PromQL Telemetry Diagnostics & Alert Triage
This dataset contains 600 curated training records with in-depth, verbose 4-phase <Thinking> Chain-of-Thought reasoning, 100 frozen evaluation benchmark samples, and 50 frozen regression verification samples formatted in standard ChatML (messages) and Prompt-Target pairs, strictly following the Pioneer / Prometheus research paper 3-slice curriculum design.
📊 Dataset Composition & 3-Slice Breakdown… See the full description on the dataset page: https://huggingface.co/datasets/StarsMakeGalaxy/devops-promql-sre-curated-600.devops-v1
Dataset Card for Docker & Kubernetes Troubleshooting Dataset
Dataset Summary
This dataset contains 256 comprehensive question-solution pairs covering Docker and Kubernetes troubleshooting scenarios. It includes common issues, error messages, and their detailed solutions for container orchestration and cloud-native infrastructure management. The dataset spans Docker fundamentals, Kubernetes core concepts, cloud-managed Kubernetes services (EKS, AKS), and advanced… See the full description on the dataset page: https://huggingface.co/datasets/pavanmantha/devops-v1.devops_augsoftware-engineering-and-devopsdevopseval-examdevops_sjDevopsproduct-dev-ops-playbook
Product Dev Ops Playbook — Cross-Functional Sprint Alignment SOP
Complete Product × Engineering × Operations alignment SOP — from unified backlog to 10-day sprint cadence to veto power rules.
📦 Install on ClawHub
clawhub install product-dev-ops-playbook
Then ask your AI agent:
"Design a 10-day sprint cadence for our 8-person product+eng team"
Installs the complete Product × Engineering × Operations alignment SOP — dual-layer Kanban, 10-day sprint cadence… See the full description on the dataset page: https://huggingface.co/datasets/Gingiris/product-dev-ops-playbook.my-issues-dataset
Dataset Card for Dataset Name
Dataset Summary in English
This customized dataset is made of a corpus of commun Github issues, typically utilized for tracking bugs or features within a repositories. This self-constructed corpus can serve multiple purposes, such as analyzing the time taken to resolve open issues or pull requests, training a classifier to tag issues based on their descriptions (e.g., "bug," "enhancement," "question"), or developing a semantic search engine… See the full description on the dataset page: https://huggingface.co/datasets/devopsmarc/my-issues-dataset.devops-gitops-corpusdevops-opsnotes-instructions
