13b
Datasets
All datasets matching “13b”conLlama-2-13b-KronQ-HG
Llama-2-13b — KronQ H_G (output-side gradient covariance)
Paper: arXiv:2607.07964 · Code: GitHub
Pre-computed H_G for Llama-2-13b, the output-side curvature factor used by KronQ under the K-FAC factorization H ≈ H_X ⊗ H_G. H_G is the per-sublayer empirical-Fisher gradient covariance (E[g gᵀ] over the layer output), distinct from the standard input-side Hessian H_X (built online during calibration).
Publishing this lets you reproduce KronQ quantization without the offline Fisher… See the full description on the dataset page: https://huggingface.co/datasets/donghyunli/Llama-2-13b-KronQ-HG.details_meta-llama__Llama-2-13b-hf
Dataset Card for Evaluation run of meta-llama/Llama-2-13b-hf
Dataset Summary
Dataset automatically created during the evaluation run of model meta-llama/Llama-2-13b-hf on the Open LLM Leaderboard.
The dataset is composed of 123 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 8 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_meta-llama__Llama-2-13b-hf.obelics_100k-tokenized-4image_llava_vicuna-13B_4096details_ajibawa-2023__OpenHermes-2.5-Code-290k-13B
Dataset Card for Evaluation run of ajibawa-2023/OpenHermes-2.5-Code-290k-13B
Dataset automatically created during the evaluation run of model ajibawa-2023/OpenHermes-2.5-Code-290k-13B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ajibawa-2023__OpenHermes-2.5-Code-290k-13B.details_huggyllama__llama-13b
Dataset Card for Evaluation run of huggyllama/llama-13b
Dataset Summary
Dataset automatically created during the evaluation run of model huggyllama/llama-13b on the Open LLM Leaderboard.
The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_huggyllama__llama-13b.
