datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DeltaSecommits_codellama-7b-hf_tokenizedbigcodebench_codellama_codellama-7b-instruct-hf_tokenizeddetails_codellama__CodeLlama-34b-hf
Dataset Card for Evaluation run of codellama/CodeLlama-34b-hf
Dataset automatically created during the evaluation run of model codellama/CodeLlama-34b-hf on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_codellama__CodeLlama-34b-hf.llama_3.1-sae-23-29-code-activationsKodCode_codellama-7b-hf_tokenizeddetails_uukuguy__speechless-codellama-platypus-13b
Dataset Card for Evaluation run of uukuguy/speechless-codellama-platypus-13b
Dataset Summary
Dataset automatically created during the evaluation run of model uukuguy/speechless-codellama-platypus-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_uukuguy__speechless-codellama-platypus-13b.details_itsliupeng__llama2_7b_code
Dataset Card for Evaluation run of itsliupeng/llama2_7b_code
Dataset Summary
Dataset automatically created during the evaluation run of model itsliupeng/llama2_7b_code on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_itsliupeng__llama2_7b_code.details_codellama__CodeLlama-7b-hf
Dataset Card for Evaluation run of codellama/CodeLlama-7b-hf
Dataset Summary
Dataset automatically created during the evaluation run of model codellama/CodeLlama-7b-hf on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_codellama__CodeLlama-7b-hf.details_codellama__CodeLlama-70b-Python-hf
Dataset Card for Evaluation run of codellama/CodeLlama-70b-Python-hf
Dataset automatically created during the evaluation run of model codellama/CodeLlama-70b-Python-hf on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_codellama__CodeLlama-70b-Python-hf.codellama-generationsHere you can find the solutions generated by of the Code Llama models to the HumanEval and multiPL-E benchmarks used in the Big Code models Leaderboard: https://huggingface.co/spaces/bigcode/bigcode-models-leaderboard.
details_TheBloke__CodeLlama-34B-Python-fp16
Dataset Card for Evaluation run of TheBloke/CodeLlama-34B-Python-fp16
Dataset Summary
Dataset automatically created during the evaluation run of model TheBloke/CodeLlama-34B-Python-fp16 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TheBloke__CodeLlama-34B-Python-fp16.details_OpenAssistant__codellama-13b-oasst-sft-v10
Dataset Card for Evaluation run of OpenAssistant/codellama-13b-oasst-sft-v10
Dataset Summary
Dataset automatically created during the evaluation run of model OpenAssistant/codellama-13b-oasst-sft-v10 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_OpenAssistant__codellama-13b-oasst-sft-v10.Llama-3.3-Future-Code-Instructions
Llama 3.3 Future Code Instructions
Llama 3.3 Future Code Instructions is a large-scale instruction dataset synthesized with the Meta Llama 3.3 70B Instruct model.
The dataset was generated with the method called Magpie, where we prompted the model to generate instructions likely to be asked by the users.
In addition to the original prompt introduced by the authors, we conditioned the system prompt on what specific programming language the user has an interest in, gaining control… See the full description on the dataset page: https://huggingface.co/datasets/future-architect/Llama-3.3-Future-Code-Instructions.details_uukuguy__speechless-codellama-orca-13b
Dataset Card for Evaluation run of uukuguy/speechless-codellama-orca-13b
Dataset Summary
Dataset automatically created during the evaluation run of model uukuguy/speechless-codellama-orca-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_uukuguy__speechless-codellama-orca-13b.details_TheBloke__CodeLlama-13B-Instruct-fp16
Dataset Card for Evaluation run of TheBloke/CodeLlama-13B-Instruct-fp16
Dataset Summary
Dataset automatically created during the evaluation run of model TheBloke/CodeLlama-13B-Instruct-fp16 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TheBloke__CodeLlama-13B-Instruct-fp16.details_uukuguy__speechless-codellama-dolphin-orca-platypus-13b
Dataset Card for Evaluation run of uukuguy/speechless-codellama-dolphin-orca-platypus-13b
Dataset Summary
Dataset automatically created during the evaluation run of model uukuguy/speechless-codellama-dolphin-orca-platypus-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_uukuguy__speechless-codellama-dolphin-orca-platypus-13b.details_ehartford__Samantha-1.11-CodeLlama-34b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-CodeLlama-34b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-CodeLlama-34b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-CodeLlama-34b.details_codellama__CodeLlama-34b-Python-hf
Dataset Card for Evaluation run of codellama/CodeLlama-34b-Python-hf
Dataset automatically created during the evaluation run of model codellama/CodeLlama-34b-Python-hf on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_codellama__CodeLlama-34b-Python-hf.details_rombodawg__test_dataset_Codellama-3-8Bdetails_lqtrung1998__Codellama-7b-hf-ReFT-Rerank-GSM8k
Dataset Card for Evaluation run of lqtrung1998/Codellama-7b-hf-ReFT-Rerank-GSM8k
Dataset automatically created during the evaluation run of model lqtrung1998/Codellama-7b-hf-ReFT-Rerank-GSM8k on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lqtrung1998__Codellama-7b-hf-ReFT-Rerank-GSM8k.details_codellama__CodeLlama-13b-hf
Dataset Card for Evaluation run of codellama/CodeLlama-13b-hf
Dataset Summary
Dataset automatically created during the evaluation run of model codellama/CodeLlama-13b-hf on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_codellama__CodeLlama-13b-hf.details_uukuguy__speechless-codellama-34b-v1.9
Dataset Card for Evaluation run of uukuguy/speechless-codellama-34b-v1.9
Dataset Summary
Dataset automatically created during the evaluation run of model uukuguy/speechless-codellama-34b-v1.9 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_uukuguy__speechless-codellama-34b-v1.9.code_contests_llamabase_mc_intermediatedetails_ehartford__WizardLM-1.0-Uncensored-CodeLlama-34b
Dataset Card for Evaluation run of ehartford/WizardLM-1.0-Uncensored-CodeLlama-34b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/WizardLM-1.0-Uncensored-CodeLlama-34b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__WizardLM-1.0-Uncensored-CodeLlama-34b.details_NousResearch__CodeLlama-34b-hf
Dataset Card for Evaluation run of NousResearch/CodeLlama-34b-hf
Dataset Summary
Dataset automatically created during the evaluation run of model NousResearch/CodeLlama-34b-hf on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_NousResearch__CodeLlama-34b-hf.cortex-codellamadetails_shareAI__CodeLLaMA-chat-13b-Chinese
Dataset Card for Evaluation run of shareAI/CodeLLaMA-chat-13b-Chinese
Dataset Summary
Dataset automatically created during the evaluation run of model shareAI/CodeLLaMA-chat-13b-Chinese on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_shareAI__CodeLLaMA-chat-13b-Chinese.details_layoric__llama-2-13b-code-alpaca
Dataset Card for Evaluation run of layoric/llama-2-13b-code-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model layoric/llama-2-13b-code-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_layoric__llama-2-13b-code-alpaca.details_ibivibiv__llama3-8b-instruct-codedetails_codellama__CodeLlama-70b-Instruct-hf
Dataset Card for Evaluation run of codellama/CodeLlama-70b-Instruct-hf
Dataset automatically created during the evaluation run of model codellama/CodeLlama-70b-Instruct-hf on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_codellama__CodeLlama-70b-Instruct-hf.
