datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qrecc
QReCC Topics
This repository hosts the QReCC topics with passage relevance.
This dataset complements the QReCC retrieval setup outlined in the Apple ML-QReCC GitHub repository.
Train split has 63501 examples, and test split 16451 examples. Relevant passages are in the field "Truth_passages".
from datasets import load_dataset
def main():
# 1. Load the dataset
dataset = load_dataset("slupart/qrecc")
# 2. Show the available splits
print("Available splits:"… See the full description on the dataset page: https://huggingface.co/datasets/slupart/qrecc.autonomous-cloud-gpu-slurm-serving-suite
⚡ Autonomous Cloud GPU Infrastructure, Slurm Orchestration & Distributed Serving Suite (2026)
A Production-Grade, Verifiable Synthetic Corpus for Training Autonomous AI Supercomputing & LLM Serving Agents
⚡ Overview & Industry Problem
Operating massive AI supercomputers (thousands of NVIDIA H100/H200 and Blackwell GPUs) requires coordinating Slurm cluster schedules, topology-aware NVLink cliques, NCCL AllReduce rings, RoCE v2 lossless fabrics… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/autonomous-cloud-gpu-slurm-serving-suite.
