CoolFace
20 results

llama2-7b

donghyunli /Llama-2-7b-KronQ-HG Llama-2-7b — KronQ H_G (output-side gradient covariance) Paper: arXiv:2607.07964 · Code: GitHub Pre-computed H_G for Llama-2-7b, the output-side curvature factor used by KronQ under the K-FAC factorization H ≈ H_X ⊗ H_G. H_G is the per-sublayer sampled-Fisher gradient covariance (labels drawn from the model distribution) (E[g gᵀ] over the layer output), distinct from the standard input-side Hessian H_X (which GPTQ/GPTAQ build online during calibration). Publishing this lets you… See the full description on the dataset page: https://huggingface.co/datasets/donghyunli/Llama-2-7b-KronQ-HG.text-generation0 likes2.7k downloads2mo agoHugging Faceautomated-research-group /llama2_7b_chat-boolq-results Dataset Card for "llama2_7b_chat-boolq-results" More Information needed text100K<n<1M1 likes2k downloads3y agoHugging Faceautomated-research-group /llama2_7b_chat-piqa-resultstext100K<n<1M0 likes1.3k downloads3y agoHugging Faceopen-llm-leaderboard-old /details_Azure99__blossom-v2-llama2-7b Dataset Card for Evaluation run of Azure99/blossom-v2-llama2-7b Dataset Summary Dataset automatically created during the evaluation run of model Azure99/blossom-v2-llama2-7b on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Azure99__blossom-v2-llama2-7b.0 likes457 downloads3y agoHugging Faceautomated-research-group /llama2_7b_chat-siqa-resultstext100K<n<1M0 likes440 downloads3y agoHugging Faceopen-llm-leaderboard-old /details_itsliupeng__llama2_7b_code Dataset Card for Evaluation run of itsliupeng/llama2_7b_code Dataset Summary Dataset automatically created during the evaluation run of model itsliupeng/llama2_7b_code on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_itsliupeng__llama2_7b_code.0 likes232 downloads3y agoHugging Face