datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
conLlama-2-13b-KronQ-HG
Llama-2-13b — KronQ H_G (output-side gradient covariance)
Paper: arXiv:2607.07964 · Code: GitHub
Pre-computed H_G for Llama-2-13b, the output-side curvature factor used by KronQ under the K-FAC factorization H ≈ H_X ⊗ H_G. H_G is the per-sublayer empirical-Fisher gradient covariance (E[g gᵀ] over the layer output), distinct from the standard input-side Hessian H_X (built online during calibration).
Publishing this lets you reproduce KronQ quantization without the offline Fisher… See the full description on the dataset page: https://huggingface.co/datasets/donghyunli/Llama-2-13b-KronQ-HG.details_meta-llama__Llama-2-13b-hf
Dataset Card for Evaluation run of meta-llama/Llama-2-13b-hf
Dataset Summary
Dataset automatically created during the evaluation run of model meta-llama/Llama-2-13b-hf on the Open LLM Leaderboard.
The dataset is composed of 123 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 8 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_meta-llama__Llama-2-13b-hf.obelics_100k-tokenized-4image_llava_vicuna-13B_4096details_ajibawa-2023__OpenHermes-2.5-Code-290k-13B
Dataset Card for Evaluation run of ajibawa-2023/OpenHermes-2.5-Code-290k-13B
Dataset automatically created during the evaluation run of model ajibawa-2023/OpenHermes-2.5-Code-290k-13B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ajibawa-2023__OpenHermes-2.5-Code-290k-13B.details_huggyllama__llama-13b
Dataset Card for Evaluation run of huggyllama/llama-13b
Dataset Summary
Dataset automatically created during the evaluation run of model huggyllama/llama-13b on the Open LLM Leaderboard.
The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_huggyllama__llama-13b.details_ceadar-ie__FinanceConnect-13B
Dataset Card for Evaluation run of ceadar-ie/FinanceConnect-13B
Dataset Summary
Dataset automatically created during the evaluation run of model ceadar-ie/FinanceConnect-13B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ceadar-ie__FinanceConnect-13B.AA_preference_vicuna-13b_l0_cutDeltaSecommits_llama-2-13b-chat_tokenized_v3_vulnerableAA_preference_vicuna-13b_cosi_cutdetails_umd-zhou-lab__claude2-alpaca-13B
Dataset Card for Evaluation run of umd-zhou-lab/claude2-alpaca-13B
Dataset automatically created during the evaluation run of model umd-zhou-lab/claude2-alpaca-13B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_umd-zhou-lab__claude2-alpaca-13B.details_PocketDoc__Dans-PileOfSets-Mk1-llama-13b-merged
Dataset Card for Evaluation run of PocketDoc/Dans-PileOfSets-Mk1-llama-13b-merged
Dataset Summary
Dataset automatically created during the evaluation run of model PocketDoc/Dans-PileOfSets-Mk1-llama-13b-merged on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_PocketDoc__Dans-PileOfSets-Mk1-llama-13b-merged.details_PocketDoc__Dans-RetroRodeo-13b
Dataset Card for Evaluation run of PocketDoc/Dans-RetroRodeo-13b
Dataset Summary
Dataset automatically created during the evaluation run of model PocketDoc/Dans-RetroRodeo-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_PocketDoc__Dans-RetroRodeo-13b.details_PocketDoc__Dans-MysteryModel-13b
Dataset Card for Evaluation run of PocketDoc/Dans-MysteryModel-13b
Dataset Summary
Dataset automatically created during the evaluation run of model PocketDoc/Dans-MysteryModel-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_PocketDoc__Dans-MysteryModel-13b.AA_preference_vicuna-13b_l0_fulldetails_ajibawa-2023__Uncensored-Jordan-13B
Dataset Card for Evaluation run of ajibawa-2023/Uncensored-Jordan-13B
Dataset Summary
Dataset automatically created during the evaluation run of model ajibawa-2023/Uncensored-Jordan-13B on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ajibawa-2023__Uncensored-Jordan-13B.details_facebook__opt-13b
Dataset Card for Evaluation run of facebook/opt-13b
Dataset Summary
Dataset automatically created during the evaluation run of model facebook/opt-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_facebook__opt-13b.AA_preference_vicuna-13b_cooccur_fullAA_preference_vicuna-13b_cosi_fulldetails_TheBloke__Llama-2-13B-GPTQ
Dataset Card for Evaluation run of TheBloke/Llama-2-13B-GPTQ
Dataset Summary
Dataset automatically created during the evaluation run of model TheBloke/Llama-2-13B-GPTQ on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TheBloke__Llama-2-13B-GPTQ.details_ehartford__WizardLM-1.0-Uncensored-Llama2-13b
Dataset Card for Evaluation run of ehartford/WizardLM-1.0-Uncensored-Llama2-13b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/WizardLM-1.0-Uncensored-Llama2-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__WizardLM-1.0-Uncensored-Llama2-13b.details_psmathur__orca_mini_v2_13b
Dataset Card for Evaluation run of psmathur/orca_mini_v2_13b
Dataset Summary
Dataset automatically created during the evaluation run of model psmathur/orca_mini_v2_13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_psmathur__orca_mini_v2_13b.pretrain-dataset-T2048-13B
Pretrain Dataset (Tokenized)
This dataset contains tokenized and packed sequences ready for LLM pretraining.
Dataset Details
Property
Value
Sequences
6,474,097
Sequence Length
2048
Tokenizer
./vn_spm_v3_fast2/
Total Tokens
13,258,950,332
Shards
13
Created
2025-12-10
Dataset Structure
Each sample contains:
input_ids: List of token IDs (length: 2048)
attention_mask: Attention mask (1 for real tokens, 0 for padding)… See the full description on the dataset page: https://huggingface.co/datasets/tvu-vlinhd11/pretrain-dataset-T2048-13B.details_uukuguy__speechless-codellama-platypus-13b
Dataset Card for Evaluation run of uukuguy/speechless-codellama-platypus-13b
Dataset Summary
Dataset automatically created during the evaluation run of model uukuguy/speechless-codellama-platypus-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_uukuguy__speechless-codellama-platypus-13b.details_posicube__Llama2-chat-AYT-13B
Dataset Card for Evaluation run of posicube/Llama2-chat-AYT-13B
Dataset Summary
Dataset automatically created during the evaluation run of model posicube/Llama2-chat-AYT-13B on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_posicube__Llama2-chat-AYT-13B.details_openlm-research__open_llama_13b
Dataset Card for Evaluation run of openlm-research/open_llama_13b
Dataset Summary
Dataset automatically created during the evaluation run of model openlm-research/open_llama_13b on the Open LLM Leaderboard.
The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_openlm-research__open_llama_13b.lm-eval-results-yunconglong-DARE_TIES_13B-private
Dataset Card for Evaluation run of yunconglong/DARE_TIES_13B
Dataset automatically created during the evaluation run of model yunconglong/DARE_TIES_13B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yunconglong-DARE_TIES_13B-private.preference-test-olmo-32b-13bAA_preference_vicuna-13b_cooccur_cutdetails_core42__jais-13b
Dataset Card for Evaluation run of core42/jais-13b
Dataset automatically created during the evaluation run of model core42/jais-13b.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_core42__jais-13b.
