datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Openthoughts_math_30k_opsdu-opsd-backupops-eval
OPS-Eval: Leakage-Resistant Evaluation for Optical Pooled Screens
Benchmark artifacts for evaluating representation learning on pooled CRISPR
microscopy data. This dataset accompanies a submission to the NeurIPS 2026
Evaluations and Datasets Track.
Contents
Directory/File
Description
Size
montages/
Per-gene montage images (4 channels x 2 phases, ~10 PNGs per gene)
~66 GB
cell_embeddings/
Pre-extracted 512-dim cell embeddings per sgRNA (.npz)
~15 GB… See the full description on the dataset page: https://huggingface.co/datasets/cspeters119/ops-eval.opsd-instruction-scale-omni-full-v1-artifacts-publicNGS
Noise Guided Splatting (NGS) Transparency Datasets
GitHub | Project Page
This repository contains the datasets used in the paper "Fix False Transparency by Noise Guided Splatting". It is designed to facilitate research and benchmarking for the "false transparency" artifact in 3D Gaussian Splatting (3DGS) reconstructions of opaque objects.
The repository is composed of four distinct subsets, each augmented with noise Gaussian infills (inside_gaussians.ply) crucial for evaluating… See the full description on the dataset page: https://huggingface.co/datasets/OpsiClear/NGS.ops-lite
ops-lite
A curated 500-case root-cause-analysis (RCA) evaluation set for
microservice systems, with manifest-driven causal-graph ground truth.
Each case bundles:
a chaos-injection ground truth (injection.json)
a causal service graph derived from the injection's fault contract
(causal_graph.json)
the runtime environment snapshot (env.json, result.json,
label.txt)
12 parquet metric tables per case, split into the abnormal window
(during fault) and the normal window (baseline)
The… See the full description on the dataset page: https://huggingface.co/datasets/anon-ops/ops-lite.EEGWORLDCUDA-Agent-Ops-6K
CUDA-Agent-Ops-6K
CUDA-Agent-Ops-6K is a curated training dataset for CUDA kernel generation and optimization.
It is released as part of the CUDA-Agent project:
Project Page: https://CUDA-Agent.github.io/
Github Repo: https://github.com/BytedTsinghua-SIA/CUDA-Agent
Dataset Summary
CUDA-Agent-Ops-6K contains 6,000 synthesized operator-level training tasks designed for large-scale agentic RL training. It is intended to provide diverse and executable CUDA-oriented training… See the full description on the dataset page: https://huggingface.co/datasets/BytedTsinghua-SIA/CUDA-Agent-Ops-6K.opsd-metaopenthoughts_math_30k_opsd_splitgit-ops-recovery-trajectories
Git Ops Recovery Trajectories
Rights & intended use: legacy public research corpus / portfolio
artifact. Hosted frontier-model outputs are research-only inputs under
project policy (synthetic-factory#161):
intended_use: research_only, project_training_policy: blocked. Not
training data for any model-weight update. Machine-readable record:
rights.json.
Release status: The raw, uncurated payload is now published under
data/raw/. It is available for inspection and… See the full description on the dataset page: https://huggingface.co/datasets/rmems/git-ops-recovery-trajectories.dc-ops-dataset
DC-Ops: Data Center Components Dataset
On-device data center operations assistant dataset for the Qualcomm x Meta ExecuTorch Hackathon.
Overview
319 images of data center infrastructure (server racks, NVIDIA NVL72, cables, ports, etc.)
3,045 polygon annotations in YOLO-seg format
16 component classes auto-labeled with Grounding DINO + SAM, for fine-tuning YOLOv8n-seg
Classes
ID
Class
Count
0
server rack
rack enclosures
1
compute tray… See the full description on the dataset page: https://huggingface.co/datasets/abhijitbetigeri/dc-ops-dataset.bitwise-opsstring-opsgsm_infinite_hard_r0.4_ops24opsd-cl-assets
opsd-cl-assets
Offline assets for the OPSD continual-learning project on a box without DNS.
Built 2026-09-12 on VAST. Pull everything with one command, then run install_on_box.sh.
path
content
size
wheelhouse/
167 wheels resolved from siyan-zhao/OPSD environment.yml for cp310 / manylinux x86_64, plus deepspeed-0.18.2.tar.gz and the lock file
4.7 GB
models/Qwen3-1.7B
Qwen/Qwen3-1.7B
3.8 GB
datasets/Openthoughts_math_30k_opsd
siyanzhao/Openthoughts_math_30k_opsd… See the full description on the dataset page: https://huggingface.co/datasets/HzChen20/opsd-cl-assets.ops-lite-review-sample
ops-lite-review-sample
This repository is a reviewer-facing sample of the full ops-lite benchmark.
Full dataset: https://huggingface.co/datasets/anon-ops/ops-lite
Full dataset release used for this sample: the same public ops-lite
release associated with this submission
Sample dataset URL: https://huggingface.co/datasets/anon-ops/ops-lite-review-sample
What is included
10 complete RCA cases under cases/
a subset manifest.jsonl containing metadata for those same 10… See the full description on the dataset page: https://huggingface.co/datasets/anon-ops/ops-lite-review-sample.adaptive-operator-v4-dataset
Adaptive Operator v4 — Training Datasets
Training data for adaptive-operator-v4, a Qwen3.5-9B model fine-tuned with custom control tokens for adaptive compute allocation in agentic workflows.
Dataset Summary
Split
Examples
Format
Size
SFT
4,992
OpenAI chat messages
8.8 MB
DPO
5,000
Chosen/rejected pairs
8.3 MB
Raw teacher responses
5,000
OpenAI chat messages
6.5 MB
Improved responses
4,992
OpenAI chat messages
13 MB
Preference pairs (raw)
5,000… See the full description on the dataset page: https://huggingface.co/datasets/davidnichols-ops/adaptive-operator-v4-dataset.gsm_infinite_hard_r0.4_ops16op-spp-streams-v2
op-spp-streams-v1 — tokenized Megatron streams (SPP-format pretraining corpus)
Tokenized Dolma v1.7 (ODC-BY)
subsample in Megatron IndexedDataset format (uint16, SmolLM2 tokenizer +
<assistant> extension from epfl-dlab/spp-training):
compact dense-packed 2049-token windows; annotated/canary one document
per window, EOD-padded. Built for a Synthetic-Persona-Pretraining-recipe run
(arXiv:2608.13482) — these files carry ONLY document tokens (raw public-corpus
text); the persona… See the full description on the dataset page: https://huggingface.co/datasets/joshycodes/op-spp-streams-v2.OpsEval
OpsEval Dataset
Website | Reporting Issues
Introduction
The OpsEval dataset represents a pioneering effort in the evaluation of Artificial Intelligence for IT Operations (AIOps), focusing on the application of Large Language Models (LLMs) within this domain. In an era where IT operations are increasingly reliant on AI technologies for automation and efficiency, understanding the performance of LLMs in operational tasks becomes crucial. OpsEval offers a comprehensive… See the full description on the dataset page: https://huggingface.co/datasets/Junetheriver/OpsEval.train-opsdcloud-ops-bench-datasetFull_Agent_RL_OPSD_with_Just_2_A800srubric-grounded-faithfulness-eval
Rubric-Grounded Faithfulness Evaluation Resource
This repository hosts the anonymized evaluation resource accompanying the NeurIPS 2026 Evaluations and Datasets submission:
From Scores to Checks: Rubric-Grounded Faithfulness Evaluation for AI-Generated Images
The release contains the derived assets behind the paper's main claims: full-gold AIGCIQA2023 rubric labels, evidence-point and reviewed-counterfactual diagnostic subsets, a relabeled 2400-image T2I-CompBench human-eval… See the full description on the dataset page: https://huggingface.co/datasets/Anonymous1-afk-ops/rubric-grounded-faithfulness-eval.gsm_infinite_hard_r0.4_ops8unlabelled-data
HTD Scraped Datasets
Cleaned text chunks from Home Team Department (HTD) agency websites, processed into a consistent 8-column format for downstream tasks (QA pairs annotation, fine-tuning, etc.)
Schema
Column
Type
Description
chunk_text
str
Cleaned text chunk (40-650 words)
source_url
str
Original page URL
page_title
str
Page title
section_headers
str
Section heading path
heading_path
str
Hierarchical path from site structure
content_type
str… See the full description on the dataset page: https://huggingface.co/datasets/ops-tuned-llm/unlabelled-data.swe_opsd_datasetnumber-ops
Number Ops
This dataset is part of the Roboflow 100 benchmark, a diverse collection of 100 object detection datasets spanning 7 imagery domains.
Dataset Statistics
Split
Images
Train
4,869
Validation
1,636
Test
623
Total
7,128
Classes (15)
0
1
2
3
4
5
6
7
8
9
div
eqv
minus
mult
plus
Usage
With LibreYOLO
from libreyolo import LIBREYOLO
# Load a model
model = LIBREYOLO(model_path="libreyoloXnano.pt")
# Train on this… See the full description on the dataset page: https://huggingface.co/datasets/LibreYOLO/number-ops.excel-ai-ops
ExcelAI Ops Synthetic
Programmatic synthetic data for training a 1B specialist to generate Excel workbooks via constrained JSON ops.
Each row: instruction + input_data (CSV) -> completion (JSON string of ops list). Parse completion with json.loads, validate with schema.json, build with build_excel.py -> .xlsx.
Tasks: budget_tracker, invoice, sales_report, gradebook, inventory, timesheet.
Format:
prompt: "Instruction: ...\nData:\n...\nOutput JSON ops:"
completion: JSON string of… See the full description on the dataset page: https://huggingface.co/datasets/maxie-12321/excel-ai-ops.
