datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mlops-deployment-sft-100k
MLOps Deployment SFT 100K
A synthetic supervised fine-tuning dataset of 100,000 high-quality conversations covering MLOps and ML model deployment — from model serving and inference optimization to monitoring, CI/CD, and production scaling. Designed to train AI assistants that can help ML engineers deploy and operate models at scale.
Dataset Description
This dataset covers the full MLOps lifecycle across 13 specialized categories. Each record follows the ShareGPT… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/mlops-deployment-sft-100k.mlops-infrastructure-fr
MLOps & Infrastructure IA - Dataset Francais
Dataset professionnel bilingue couvrant le MLOps, l'infrastructure GPU, la securite des pipelines ML, l'optimisation des couts LLM et les bases vectorielles en production.
Description
Ce dataset a ete cree par AYI & NEDJIMI Consultants pour fournir des ressources structurees et professionnelles sur le deploiement de systemes d'IA en production. Il couvre l'ensemble de la chaine MLOps, de l'experimentation au serving, en… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/mlops-infrastructure-fr.mlops-infrastructure-en
MLOps & AI Infrastructure - English Dataset
Professional bilingual dataset covering MLOps, GPU infrastructure, ML pipeline security, LLM cost optimization, and vector databases in production.
Description
This dataset was created by AYI & NEDJIMI Consultants to provide structured and professional resources on deploying AI systems in production. It covers the entire MLOps chain, from experimentation to serving, including security and cost optimization.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/mlops-infrastructure-en.mlops-agent-rl-eval
MLOps Agent RL — Evaluation Results
Offline multi-turn tool-calling evaluation on the mocked MLOpsEnv
incident suite (6 scenarios).
Models
base: Qwen/Qwen2.5-0.5B-Instruct
grpo: vanishingradient/mlops-agent-rl-qwen05b-grpo (LoRA adapter)
oracle: scripted expected tools + preferred remediation
Files
episodes.jsonl — full trajectories + metrics
episodes_slim.jsonl — metrics without trajectories
summary.json — aggregates for the paper
config.json —… See the full description on the dataset page: https://huggingface.co/datasets/vanishingradient/mlops-agent-rl-eval.
