datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
laion2b-en-a65_cogvlm2-4bit_captions
Abstract
This dataset contains image captions for the laion2B-en aesthetics>=6.5 image dataset using CogVLM2-4bit with the "laion-pop"-prompt to generate captions which were "likely" used in Stable Diffusion 3 training. From these image captions new synthetic images were generated using stable-diffusion-3-medium (batch-size=8).
The synthetic images are best viewed locally by cloning this repo with:
git lfs install
git clone… See the full description on the dataset page: https://huggingface.co/datasets/GeroldMeisinger/laion2b-en-a65_cogvlm2-4bit_captions.details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v12
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v12.Bitext-SmolLM2-1024-natural-instructions-formatQwen3.5-27B-AWQ-4bit-GPQA-Diamond-benchmarkBenchmark of cyankiwi/Qwen3.5-27B-AWQ-4bit against fingertap/GPQA-Diamond dataset.
Accuracy: 76.3% with Python tool.
Metric
Value
Correct
151
Incorrect
46
Errors
1
Total samples
198
Python tool calls
225
Total completion tokens
659,879
Raw stats:
{
"accuracy": 0.763,
"correct": 151,
"incorrect": 46,
"error": 1,
"total": 198,
"python_tool_calls": 225,
"completion_tokens": 659879
}
details_robinsmits__Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit
Dataset Card for Evaluation run of robinsmits/Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit
Dataset automatically created during the evaluation run of model robinsmits/Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_robinsmits__Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit.details_cloudyu__4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO
Dataset Card for Evaluation run of cloudyu/4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO
Dataset automatically created during the evaluation run of model cloudyu/4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO.details_Enno-Ai__vigogne2-enno-13b-sft-lora-4bit
Dataset Card for Evaluation run of Enno-Ai/vigogne2-enno-13b-sft-lora-4bit
Dataset Summary
Dataset automatically created during the evaluation run of model Enno-Ai/vigogne2-enno-13b-sft-lora-4bit on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Enno-Ai__vigogne2-enno-13b-sft-lora-4bit.Deepfake-and-real-images-4details_TFLai__llama-2-13b-4bit-alpaca-gpt4
Dataset Card for Evaluation run of TFLai/llama-2-13b-4bit-alpaca-gpt4
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/llama-2-13b-4bit-alpaca-gpt4 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-2-13b-4bit-alpaca-gpt4.white_paper_holdout_4___RealVisXL_V4.0details_TFLai__llama-13b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/llama-13b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/llama-13b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-13b-4bit-alpaca.details_TFLai__llama-7b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/llama-7b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/llama-7b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-7b-4bit-alpaca.sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522-details
Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522
Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522-details.details_TFLai__gpt-neox-20b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/gpt-neox-20b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/gpt-neox-20b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__gpt-neox-20b-4bit-alpaca.details_LimYeri__CodeMind-Gemma-7B-QLoRA-4bit
Dataset Card for Evaluation run of LimYeri/CodeMind-Gemma-7B-QLoRA-4bit
Dataset automatically created during the evaluation run of model LimYeri/CodeMind-Gemma-7B-QLoRA-4bit on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LimYeri__CodeMind-Gemma-7B-QLoRA-4bit.white_paper_holdout_4___FLUX.1-devdetails_Ramikan-BR__tinyllama-coder-py-4bit-v4details_TFLai__pythia-2.8b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/pythia-2.8b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/pythia-2.8b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__pythia-2.8b-4bit-alpaca.details_alnrg2arg__test3_sft_4bit
Dataset Card for Evaluation run of alnrg2arg/test3_sft_4bit
Dataset automatically created during the evaluation run of model alnrg2arg/test3_sft_4bit on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_alnrg2arg__test3_sft_4bit.white_paper_holdout_4___stable-diffusion-xl-base-1.0details_TFLai__gpt-neo-1.3B-4bit-alpaca
Dataset Card for Evaluation run of TFLai/gpt-neo-1.3B-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/gpt-neo-1.3B-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__gpt-neo-1.3B-4bit-alpaca.details_Ramikan-BR__tinyllama-coder-py-4bit-v10details_FairMind__Phi-3-mini-4k-instruct-bnb-4bit-Itadetails_FairMind__Llama-3-8B-4bit-UltraChat-Itasonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180stepsdetails_malhajar__Platypus2-70B-instruct-4bit-gptq
Dataset Card for Evaluation run of malhajar/Platypus2-70B-instruct-4bit-gptq
Dataset Summary
Dataset automatically created during the evaluation run of model malhajar/Platypus2-70B-instruct-4bit-gptq on the Open LLM Leaderboard.
The dataset is composed of 61 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_malhajar__Platypus2-70B-instruct-4bit-gptq.details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v5sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbo-180steps-details
Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbo-180steps
Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbo-180steps
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbo-180steps-details.details_Ramikan-BR__tinyllama-coder-py-4bit-v3details_unsloth__llama-3-8b-bnb-4bit
Dataset Card for Evaluation run of unsloth/llama-3-8b-bnb-4bit
Dataset automatically created during the evaluation run of model unsloth/llama-3-8b-bnb-4bit.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_unsloth__llama-3-8b-bnb-4bit.
