Workload
Datasets
All datasets matching “Workload”Claude-opus-5-xhigh-workload-agent-preview
Overview
Vietnamese multi-turn tool-use conversations with a <think> block on every assistant turn.
Notes: this only the preview version not fully dataset
examples
368
assistant turns
842 — 100% carry <think>
reasoning generated by
claude-opus-5
format
OpenAI-chat JSONL
Configs
from datasets import load_dataset
ds = load_dataset("beyoru/misa-agentwork-reasoning") # with <think>
ds =… See the full description on the dataset page: https://huggingface.co/datasets/beyoru/Claude-opus-5-xhigh-workload-agent-preview.synthetic-cache-workloads
Synthetic Cache Workloads (Oracle Traces)
1. Dataset Overview
This dataset contains synthetic memory access traces generated to facilitate research in Machine Learning for Systems (SysML), specifically for Cache Replacement Policies.
Unlike standard raw trace logs (which only contain a list of accessed addresses), this dataset is pre-processed with Feature Engineering and Oracle Labels. It is designed to train Supervised Learning models or Reinforcement Learning agents to… See the full description on the dataset page: https://huggingface.co/datasets/rajaykumar12959/synthetic-cache-workloads.emgena_cloud_gcp_workload_identity_federation_triage_teaser
🔬 INSPECT THE DEEPSEEK-R1 REASONING CHAIN LIVE:
Zero hallucinations. Null syntax errors. 100% AST compiler validated.🌐 Live Interactive Reasoning & Code Inspector: https://emgena.com/trainingslager🎁 Claim your Free Starter Kit (Code: STARTER100): https://emgena.com/trainingslager🏷️ Launch Discount: Get 20 € OFF any 500-incident production suite with code LAUNCH20!
📜 Enterprise Compliance: EU AI Act Articles 50 & 53 certified • 100% DSGVO / GDPR clean • Commercial EULA… See the full description on the dataset page: https://huggingface.co/datasets/emgena/emgena_cloud_gcp_workload_identity_federation_triage_teaser.benchmark-dataset-different-gpu-workload
GPU catalog × LLM workload VRAM benchmark
Summary
Tabular benchmark in CSV form: each row pairs a catalog GPU (gpu_id, gpu_display_name, catalog_gpu_vram_gb) with a concrete LLM inference-style workload (model, parameter count, context length, precision, batch size, concurrent users). The file records math_engine VRAM component estimates (weights, KV cache, activations, overhead, totals, tier), a document_engine recommended VRAM value, a short comparison summary… See the full description on the dataset page: https://huggingface.co/datasets/odyn-network/benchmark-dataset-different-gpu-workload.h200-power-workload-traces-10MHz_5s
H200 GPU Power-Draw Workload Traces (10 MHz, 5 s)
Off-chip power side-channel recordings of an NVIDIA H200 NVL, captured with an
external current probe, for AI compute governance: determining
whether a GPU is training, running inference, or executing non-AI computation
from a signal that an external verifier could plausibly observe off-chip, and
how robust is detection to an operator who tries to disguise training as inference.
This is the dataset used for the paper Workload… See the full description on the dataset page: https://huggingface.co/datasets/simgar/h200-power-workload-traces-10MHz_5s.clustercast-sample-workloads
ClusterCast Sample Workloads
Synthetic GPU cluster scenarios for benchmarking schedulers.
Samples
File
Scenario
Stress
sample_ai_training.json
AI Training Cluster
Baseline
sample_inference_serving.json
Inference Serving
GPU Shortage
sample_render_farm.json
Render Farm
Training Burst
sample_scientific_compute.json
Scientific HPC
Deadline Pressure
sample_cloud_provider.json
Cloud Provider
Network Congestion
Each sample includes cluster nodes… See the full description on the dataset page: https://huggingface.co/datasets/alirezaaminzadeh/clustercast-sample-workloads.
