datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
devopsbench-100
DevOpsBench-100
DevOpsBench-100 is a synthetic long-horizon software-engineering / SRE agent
benchmark: 100 tasks over one executable world ("NovaCart", a mid-size
e-commerce SaaS) with 72 SQLite tables,
1451 seeded rows, a 38-file monorepo with 417 commits,
and 97 MCP tools spanning a first-party engineering stack
(tickets, PRs, CI, deployments, canaries, migrations, feature flags, metrics,
alerts, incidents, chat, knowledge base) plus deliberately disagreeing
vendor-shaped… See the full description on the dataset page: https://huggingface.co/datasets/SamuelChien821/devopsbench-100.stack-v3-devops
The Stack v3 DevOps Corpus
13,234,862 complete infrastructure units extracted from
The Stack v3,
grouped into seven classes and gated on content rather than popularity.
A unit is not a file, it is the thing an engineer would actually run: a Helm chart
arrives with its Chart.yaml, values.yaml and every template; a Terraform module
with all of its .tf files; an Ansible role with its tasks, defaults and handlers.
That is only possible because The Stack v3 groups rows by repository… See the full description on the dataset page: https://huggingface.co/datasets/Helmcode/stack-v3-devops.devops-incident-response
Dataset Card for DevOps Incident Response Dataset
Dataset Description
Dataset Summary
The DevOps Incident Response Dataset is a comprehensive collection of real-world-style DevOps incidents, troubleshooting scenarios, and resolution procedures. This dataset is designed to help train AI models for DevOps assistance, incident response automation, and technical troubleshooting education.
Each incident includes:
Detailed incident description and symptoms… See the full description on the dataset page: https://huggingface.co/datasets/Snaseem2026/devops-incident-response.devops-cloud-instruction-dataset
DevOps & Cloud Infrastructure Dataset
Professional instruction-response pairs for DevOps engineers covering Kubernetes, Docker, Terraform, CI/CD, and cloud services (AWS, Azure).
Dataset Details
Dataset Description
This is a high-quality instruction-tuning dataset focused on Devops Cloud topics. Each entry includes:
A clear instruction/question
Optional input context
A detailed response/solution
Chain-of-thought reasoning process
Curated by: CloudKernel.IO… See the full description on the dataset page: https://huggingface.co/datasets/bernabepuente/devops-cloud-instruction-dataset.devops-kubernetes-sft-100k
DevOps and Kubernetes SFT 100K
A synthetic supervised fine-tuning dataset of 100,000 high-quality DevOps and Kubernetes conversations designed to train AI assistants capable of supporting platform engineers, SREs, and DevOps practitioners.
Dataset Description
This dataset covers production-grade Kubernetes operations, cloud infrastructure, CI/CD pipelines, GitOps workflows, and platform engineering across 13 specialized categories. Each record follows the ShareGPT… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/devops-kubernetes-sft-100k.devops-sft-dataset
DevOps SFT Instruction Dataset
This dataset contains 8,076 high-quality instruction-response pairs specifically generated for fine-tuning a DevOps domain-specialized language model. It was used in the Supervised Fine-Tuning (SFT) phase of the Ulysses model training pipeline.
Dataset Description
Instructions were generated using the Gemini API (gemini-2.0-flash) and Ollama (qwen2.5-coder:7b) by feeding chunks of official DevOps documentation and GitHub repositories… See the full description on the dataset page: https://huggingface.co/datasets/jalpan04/devops-sft-dataset.Devops_LLMdevops-kubernetes-iac-sft-dpo-2026
⚙️ Enterprise DevOps AI, Kubernetes SRE & IaC SFT/DPO Dataset (2026)
High-precision multi-turn instruction tuning and preference optimization dataset with step-by-step SRE root-cause Chain-of-Thought (<thought>) diagnostic trees for fine-tuning LLMs (Llama-3.3, Qwen-2.5-Coder, DeepSeek-R1-Distill, Mistral) into Senior Site Reliability Engineers (SRE), Principal Cloud Architects, and DevSecOps Specialists.
📊 Dataset Architecture & Highlights
Multi-Turn SRE… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/devops-kubernetes-iac-sft-dpo-2026.devops-predictive-logs
🔮 DevOps Predictive Logs Dataset
A synthetic dataset of realistic DevOps log sequences for training and benchmarking predictive failure models.
📊 Dataset Summary
This dataset contains realistic DevOps log scenarios covering common infrastructure failure patterns. Each log entry includes metadata about the failure scenario, severity, and time-to-failure, making it ideal for training predictive models.
Total Logs: ~150+ entriesScenarios: 10 unique failure patternsFormat:… See the full description on the dataset page: https://huggingface.co/datasets/Snaseem2026/devops-predictive-logs.deepfabric-devops-reasoning-traces
Dataset Description
DevOps reasoning traces
Dataset Details
Created by: Always Further
License: CC BY 4.0
Language(s): [English
Dataset Size: 10050
Data Splits
[train]
Dataset Creation
This dataset was created using DeepFabric, an open-source tool for generating high-quality training datasets for AI models.
Usage
from datasets import load_dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/nolabs/deepfabric-devops-reasoning-traces.SO-Python_QA-System_Administration_and_DevOps_classsmoltrace-devops-tasks
SMOLTRACE Synthetic Dataset
This dataset was generated using the TraceMind MCP Server's synthetic data generation tools.
Dataset Info
Tasks: 100
Format: SMOLTRACE evaluation format
Generated: AI-powered synthetic task generation
Usage with SMOLTRACE
from datasets import load_dataset
# Load dataset
dataset = load_dataset("MCP-1st-Birthday/smoltrace-devops-tasks")
# Use with SMOLTRACE
# smoltrace-eval --model openai/gpt-4 --dataset-name… See the full description on the dataset page: https://huggingface.co/datasets/MCP-1st-Birthday/smoltrace-devops-tasks.devops-dataset-v1bio-devops-synthetic-instructions
Bio-DevOps Synthetic Instructions
This dataset contains synthetic instruction-following examples for biomedical-style data-engineering and scientific-computing workflows.
It was created for educational and portfolio use as part of a LoRA/QLoRA fine-tuning project using Qwen/Qwen2.5-Coder-7B-Instruct.
Related model:
AiLLMBS/qwen25-coder-bio-devops-lora
Dataset Contents
The dataset includes synthetic examples for:
Python CSV validation
pandas duplicate checks
bash… See the full description on the dataset page: https://huggingface.co/datasets/AiLLMBS/bio-devops-synthetic-instructions.devops-promql-sre-curated-600
🚀 DevOps SRE PromQL Telemetry Diagnostics & Alert Triage
This dataset contains 600 curated training records with in-depth, verbose 4-phase <Thinking> Chain-of-Thought reasoning, 100 frozen evaluation benchmark samples, and 50 frozen regression verification samples formatted in standard ChatML (messages) and Prompt-Target pairs, strictly following the Pioneer / Prometheus research paper 3-slice curriculum design.
📊 Dataset Composition & 3-Slice Breakdown… See the full description on the dataset page: https://huggingface.co/datasets/StarsMakeGalaxy/devops-promql-sre-curated-600.stackexchange_devopsdevops-v1
Dataset Card for Docker & Kubernetes Troubleshooting Dataset
Dataset Summary
This dataset contains 256 comprehensive question-solution pairs covering Docker and Kubernetes troubleshooting scenarios. It includes common issues, error messages, and their detailed solutions for container orchestration and cloud-native infrastructure management. The dataset spans Docker fundamentals, Kubernetes core concepts, cloud-managed Kubernetes services (EKS, AKS), and advanced… See the full description on the dataset page: https://huggingface.co/datasets/pavanmantha/devops-v1.devops_augsoftware-engineering-and-devopsdevops_sjDevopsdevops-gitops-corpusmy-issues-dataset
Dataset Card for Dataset Name
Dataset Summary in English
This customized dataset is made of a corpus of commun Github issues, typically utilized for tracking bugs or features within a repositories. This self-constructed corpus can serve multiple purposes, such as analyzing the time taken to resolve open issues or pull requests, training a classifier to tag issues based on their descriptions (e.g., "bug," "enhancement," "question"), or developing a semantic search engine… See the full description on the dataset page: https://huggingface.co/datasets/devopsmarc/my-issues-dataset.devops-opsnotes-instructionsPL-DevOps-Instructexp_8_8_domain_shift_devops_test25exp_8_8_domain_shift_devops_test5devops-training
Dataset Description
DevOps reasoning traces
Dataset Details
Created by: Always Further
License: CC BY 4.0
Language(s): [English
Dataset Size: 10050
Data Splits
[train]
Dataset Creation
This dataset was created using DeepFabric, an open-source tool for generating high-quality training datasets for AI models.
Usage
from datasets import load_dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/shahiddanial3545/devops-training.TebyanFarsidevops-guide-demo
