datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dgx-spark-benchmarks
DGX Spark LLM Arena benchmarks
Reproducible LLM inference benchmarks on an NVIDIA DGX Spark (GB10, 128 GB unified memory). The suite defines eleven tests: six closed-loop (llama-benchy) and five open-loop (vllm bench serve). Results cover all eleven: the ten throughput tests under results, and the rate sweep under rateSweep. Raw results remain inspectable, but only complete runs without a failed sanity check count toward rankings and aggregate throughput. Open-loop tests must… See the full description on the dataset page: https://huggingface.co/datasets/Djangodevreng/dgx-spark-benchmarks.dgx-spark-eval
DGX Spark Model Evaluations
75 Messläufe in fünf Konfigurationen, alle auf einer Maschine gemessen.
Keine Herstellerangaben — jede Zahl stammt aus einem eigenen Lauf. Stand: 2026-08-17.
Die Website zu denselben Daten: https://results.southbyte.de/
Was gemessen wurde
Config
Zeilen
Inhalt
llm_local
20
Sprachmodelle, lokal mit vLLM serviert
llm_saas
28
dieselben Testfälle gegen Frontier-APIs, als Referenzrahmen
guardrails
5
Guard-Modelle gegen einen… See the full description on the dataset page: https://huggingface.co/datasets/SouthByte/dgx-spark-eval.
