datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bunnycore__HyperLlama-3.1-8B-details
Dataset Card for Evaluation run of bunnycore/HyperLlama-3.1-8B
Dataset automatically created during the evaluation run of model bunnycore/HyperLlama-3.1-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__HyperLlama-3.1-8B-details.hypergraph_openthoughts30k
Hypergraph OpenThoughts Math 30K
Reasoning hypergraphs generated for the 29,434 examples in
siyanzhao/Openthoughts_math_30k_opsd.
Generation
Model: Qwen/Qwen3.6-35B-A3B-FP8
Thinking mode: disabled
Construction: semantic-step segmentation followed by primary-support DAG induction
Graph constraint: at most one earlier-step parent per semantic step
Processing order: source dataset row order
Schema
Each JSONL record contains:
row_index: source… See the full description on the dataset page: https://huggingface.co/datasets/dvtiendat/hypergraph_openthoughts30k.CultriX__Qwen2.5-14B-Hyperionv4-details
Dataset Card for Evaluation run of CultriX/Qwen2.5-14B-Hyperionv4
Dataset automatically created during the evaluation run of model CultriX/Qwen2.5-14B-Hyperionv4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__Qwen2.5-14B-Hyperionv4-details.winogeneratedhyperliquid-l4-dataHyperSwitch-Repo-CPT-Dataset-v2
Hyperswitch Rust Codebase Dataset
A comprehensive dataset extracted from the Hyperswitch open-source payment processing platform, containing 16,731 code samples across 37 modules with 6.99M tokens for training Rust code understanding and generation models.
📊 Dataset Overview
This dataset provides both file-level and granular code samples from Hyperswitch, a modern payment switch written in Rust. It's designed for training code models to understand payment processing… See the full description on the dataset page: https://huggingface.co/datasets/AdityaNarayan/HyperSwitch-Repo-CPT-Dataset-v2.hivemind-eval-benchmark
HivemindEval Compliance-Finding Benchmark — public 68-item subset
A stratified public subset of a frozen, contamination-gated benchmark for scoring the
quality of compliance findings across six UK/EU regulatory frameworks (PSD2 SCA-RTS,
NHS DSPT + UK GDPR, MOD JSP 440, Cyber Essentials Plus, DORA, EU AI Act — plus adjacent
instruments). Built and used to evaluate
Hypereum/HivemindEval; ships with
per-item gold and the raw per-item predictions of all six benchmarked models, so… See the full description on the dataset page: https://huggingface.co/datasets/Hypereum/hivemind-eval-benchmark.HyperSwitch-Repo-CPT-Dataset
Hyperswitch Rust Codebase Dataset
A comprehensive dataset extracted from the Hyperswitch open-source payment processing platform, containing 16,731 code samples across 37 modules with 6.99M tokens for training Rust code understanding and generation models.
📊 Dataset Overview
This dataset provides both file-level and granular code samples from Hyperswitch, a modern payment switch written in Rust. It's designed for training code models to understand payment processing… See the full description on the dataset page: https://huggingface.co/datasets/AdityaNarayan/HyperSwitch-Repo-CPT-Dataset.CultriX__Qwen2.5-14B-HyperMarck-dl-details
Dataset Card for Evaluation run of CultriX/Qwen2.5-14B-HyperMarck-dl
Dataset automatically created during the evaluation run of model CultriX/Qwen2.5-14B-HyperMarck-dl
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__Qwen2.5-14B-HyperMarck-dl-details.CultriX__Qwen2.5-14B-Hyper-details
Dataset Card for Evaluation run of CultriX/Qwen2.5-14B-Hyper
Dataset automatically created during the evaluation run of model CultriX/Qwen2.5-14B-Hyper
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__Qwen2.5-14B-Hyper-details.CultriX__Qwen2.5-14B-Hyperionv5-details
Dataset Card for Evaluation run of CultriX/Qwen2.5-14B-Hyperionv5
Dataset automatically created during the evaluation run of model CultriX/Qwen2.5-14B-Hyperionv5
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__Qwen2.5-14B-Hyperionv5-details.CultriX__Qwen2.5-14B-Hyperionv3-details
Dataset Card for Evaluation run of CultriX/Qwen2.5-14B-Hyperionv3
Dataset automatically created during the evaluation run of model CultriX/Qwen2.5-14B-Hyperionv3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__Qwen2.5-14B-Hyperionv3-details.
