datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
rtx-5090-benchmarks
RTX 5090 LLM Benchmarks
Speed and quality benchmarks for quantized LLMs on NVIDIA RTX 5090 32GB, measured with llm-bench-rig.
Quality Benchmarks
Generative evaluation through llama-server chat completions. Replicates standard benchmark methodology using custom evaluators — no lm-evaluation-harness dependency.
Results are split by reasoning mode: comparing a thinking-on (reasoning) model's quality against a thinking-off model is apples-to-oranges, so the two groups… See the full description on the dataset page: https://huggingface.co/datasets/omegaprime669/rtx-5090-benchmarks.Sovereign-Omega-SFT-V1
🏛️ Sovereign-Omega-SFT-V1
Overview
Sovereign-Omega-SFT-V1 is a high-fidelity synthetic reasoning dataset generated by the
Sovereign Omega Synthetic Data Factory — an autonomous pipeline built on smolagents
that generates, verifies, and recursively refines Chain-of-Thought (CoT) reasoning traces.
Key Statistics
Metric
Value
Total Trajectories
3
Verified (train split)
3
Verification Rate
100.0%
Average Confidence
0.783
Average CoT Length
527… See the full description on the dataset page: https://huggingface.co/datasets/moro72842/Sovereign-Omega-SFT-V1.acornlib-benchmark
Acorn Theorem Proving Benchmark (Preview)
Disclaimer: This is an early-stage, minimal benchmark for internal experimentation. It is not ready for academic publication or production evaluation. The task selection, difficulty calibration, and evaluation methodology are all preliminary.
A small benchmark of 50 theorems from the Acorn proof language, spanning easy to very hard. Each task asks the model to generate a valid proof body that the Acorn verifier accepts.
Data… See the full description on the dataset page: https://huggingface.co/datasets/OmegaCombinator/acornlib-benchmark.
