datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fast-autoregressive-inference-gp-trainK4sovereign-shadow-inference-bench
Sovereign Shadow Inference Bench
A public, versioned evidence surface for independent Hugging Face shadow inference beside Sovereign's primary OpenRouter/Revolver route.
What this dataset proves
The seed record in data/shadow_receipts.jsonl was produced by one real Hugging Face Inference Providers request. It records provider/model identity, request bounds, latency, hashes, literal-match outcome, source revision, and an immutable receipt hash.
What it… See the full description on the dataset page: https://huggingface.co/datasets/Thorsu/sovereign-shadow-inference-bench.qwen3.8-27b-inference-benchmark-4090
Qwen3.8-27B Inference Benchmark on RTX 4090 48GB
中文说明 · GitHub benchmark repository
Structured performance and accuracy results for four real Qwen3.8-27B serving configurations on an NVIDIA RTX 4090 48 GB workstation. A dual-GPU llama.cpp BF16 reference additionally used an RTX 3090 24 GB.
This dataset is the analysis-friendly companion to the full benchmark repository. It publishes aggregate tables, 140 normalized per-request performance records, accuracy scores, sanitized… See the full description on the dataset page: https://huggingface.co/datasets/pxzleo/qwen3.8-27b-inference-benchmark-4090.cs2-action-inference-test
CS2 战术 Action 推理测试集
本测试集用于 WAN I2V 的战术动作定性测试。每个小类只保留 1 张真实比赛 POV 第一帧,以及两种英文文本条件;本版不提供 GT 视频。第一帧来源依据 parse-dem 的 events.csv、game_events.csv 或逐 tick 状态对齐到 opencs2_matches* 视频。
数据约定
共 45 个 case、9 个大类。
每个 case 只有一张 832x480 的 first_frame.png,作为 WAN I2V 条件图;不裁剪或复制 GT clip。首帧优先选择正常持械、水平视角、无遮挡且较开阔的画面。
prompt.txt 是完整英文 prompt,包含首帧可见环境、初始持械状态、画面保持要求和整段唯一动作变化。
chunk_prompts.json 固定包含 5 个英文 prompt,依次描述期望生成视频的 0-1、1-2、2-3、3-4、4-5 秒。
metadata.json… See the full description on the dataset page: https://huggingface.co/datasets/mikusama99/cs2-action-inference-test.fast-autoregressive-inference-gp-trainK16inference-audit
Inference Audit: Provider Delivery and Metering
This dataset contains 3,932 controlled observations from OpenAI-compatible endpoints
serving openai/gpt-oss-120b through 18 pinned providers. The runs measure what an API returned
and reported at the HTTP boundary: delivery, parameter compliance, token accounting, caching,
streaming behavior, latency, and repeatability.
The records do not identify a model from its outputs, prove billing fraud, or establish why
two endpoints differ.… See the full description on the dataset page: https://huggingface.co/datasets/nuckcrews/inference-audit.enterprise-llm-inference-benchmarks-2026
🚀 Enterprise LLM Inference & Fine-Tuning Benchmarks (2026 Guide)
A curated benchmark index and architectural guide evaluating open-source foundation models, real-time inference engines (vLLM vs. TensorRT-LLM), and cloud GPU economics for enterprise deployments.
🧠 Open-Source Foundation Model Benchmarks (RAG & Code Generation)
Flagship Evaluation: Top Open-Source LLMs for Enterprise RAG & Code Generation (2026 In-Depth Guide) — Comparing Qwen 2.5 Coder, Llama… See the full description on the dataset page: https://huggingface.co/datasets/Abdulrahmankalil/enterprise-llm-inference-benchmarks-2026.fast-autoregressive-inference-scm-train5gbarxiv-author-affiliation-extraction-inference-inputs-metadatafast-autoregressive-inference-eegrepro-efficient-inference-for-noisy-llm-as-a-judge-evaluation-traces
Agent traces
Agent sessions published from a Trackio Logbook.
every-eval-ever-demostrix-halo-inference-bench
Strix Halo Local Inference Benchmarks
Measured prefill and decode throughput, and real VRAM cost, for local GGUF models
on AMD Strix Halo (Radeon 8060S / gfx1151) under ROCm.
Why this exists
Strix Halo inverts the usual local-inference trade-off. A discrete 24 GB card gives
you high memory bandwidth and a hard capacity ceiling; Strix Halo gives you the
opposite — up to 64 GiB addressable as VRAM out of 128 GB unified, at substantially
lower bandwidth. That changes… See the full description on the dataset page: https://huggingface.co/datasets/axjns/strix-halo-inference-bench.fast-autoregressive-inference-gp-trainK8laguna-xs-ultrachat-responsesinference-provider-pricinglaguna-xs-magpie-300k-responsesfast-autoregressive-inference-gp-testKFTrack_Inferencepmsp_inference_dataset_old_02pmsp_inference_dataset_old_04pmsp_inference_dataset_old_03pmsp_inference_dataset_old_01pmsp_inference_test1pmsp_inference_test10pmsp_inference_test_5_15_220
