datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
meta-llama-Llama-3.2-1B-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model.
The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl).
llama-3.2-3B-f1-instruct-eval-logs-and-scoresLlama-3.2-3B-Instruct-eval-logs-and-scoresakhadangi__Llama3.2.1B.0.01-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-First-details.NousResearch__Hermes-3-Llama-3.2-3B-details
Dataset Card for Evaluation run of NousResearch/Hermes-3-Llama-3.2-3B
Dataset automatically created during the evaluation run of model NousResearch/Hermes-3-Llama-3.2-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Hermes-3-Llama-3.2-3B-details.akhadangi__Llama3.2.1B.0.01-Last-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-Last
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-Last
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-Last-details.vhab10__Llama-3.2-Instruct-3B-TIES-details
Dataset Card for Evaluation run of vhab10/Llama-3.2-Instruct-3B-TIES
Dataset automatically created during the evaluation run of model vhab10/Llama-3.2-Instruct-3B-TIES
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/vhab10__Llama-3.2-Instruct-3B-TIES-details.meta-llama__Llama-3.2-3B-Instruct-details
Dataset Card for Evaluation run of meta-llama/Llama-3.2-3B-Instruct
Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-3B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/meta-llama__Llama-3.2-3B-Instruct-details.oopere__pruned40-llama-3.2-1B-details
Dataset Card for Evaluation run of oopere/pruned40-llama-3.2-1B
Dataset automatically created during the evaluation run of model oopere/pruned40-llama-3.2-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/oopere__pruned40-llama-3.2-1B-details.agentlans__Llama-3.2-1B-Instruct-CrashCourse12K-details
Dataset Card for Evaluation run of agentlans/Llama-3.2-1B-Instruct-CrashCourse12K
Dataset automatically created during the evaluation run of model agentlans/Llama-3.2-1B-Instruct-CrashCourse12K
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/agentlans__Llama-3.2-1B-Instruct-CrashCourse12K-details.prithivMLmods__Llama-3.2-3B-Math-Oct-details
Dataset Card for Evaluation run of prithivMLmods/Llama-3.2-3B-Math-Oct
Dataset automatically created during the evaluation run of model prithivMLmods/Llama-3.2-3B-Math-Oct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Llama-3.2-3B-Math-Oct-details.meta-llama__Llama-3.2-1B-Instruct-details
Dataset Card for Evaluation run of meta-llama/Llama-3.2-1B-Instruct
Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-1B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/meta-llama__Llama-3.2-1B-Instruct-details.Lyte__Llama-3.2-3B-Overthinker-details
Dataset Card for Evaluation run of Lyte/Llama-3.2-3B-Overthinker
Dataset automatically created during the evaluation run of model Lyte/Llama-3.2-3B-Overthinker
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lyte__Llama-3.2-3B-Overthinker-details.EpistemeAI__Reasoning-Llama-3.2-3B-Math-Instruct-RE1-details
Dataset Card for Evaluation run of EpistemeAI/Reasoning-Llama-3.2-3B-Math-Instruct-RE1
Dataset automatically created during the evaluation run of model EpistemeAI/Reasoning-Llama-3.2-3B-Math-Instruct-RE1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Reasoning-Llama-3.2-3B-Math-Instruct-RE1-details.bunnycore__Llama-3.2-3B-Booval-details
Dataset Card for Evaluation run of bunnycore/Llama-3.2-3B-Booval
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.2-3B-Booval
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.2-3B-Booval-details.cognitivecomputations__Dolphin3.0-Llama3.2-1B-details
Dataset Card for Evaluation run of cognitivecomputations/Dolphin3.0-Llama3.2-1B
Dataset automatically created during the evaluation run of model cognitivecomputations/Dolphin3.0-Llama3.2-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/cognitivecomputations__Dolphin3.0-Llama3.2-1B-details.MoonRide__Llama-3.2-3B-Khelavaster-details
Dataset Card for Evaluation run of MoonRide/Llama-3.2-3B-Khelavaster
Dataset automatically created during the evaluation run of model MoonRide/Llama-3.2-3B-Khelavaster
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/MoonRide__Llama-3.2-3B-Khelavaster-details.EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO-details
Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO-details.Nexesenex__Llama_3.2_1b_AquaSyn_0.1-details
Dataset Card for Evaluation run of Nexesenex/Llama_3.2_1b_AquaSyn_0.1
Dataset automatically created during the evaluation run of model Nexesenex/Llama_3.2_1b_AquaSyn_0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__Llama_3.2_1b_AquaSyn_0.1-details.meta-llama__Llama-3.2-3B-details
Dataset Card for Evaluation run of meta-llama/Llama-3.2-3B
Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/meta-llama__Llama-3.2-3B-details.akhadangi__Llama3.2.1B.0.1-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.1-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.1-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.1-First-details.Nexesenex__Llama_3.2_1b_Syneridol_0.2-details
Dataset Card for Evaluation run of Nexesenex/Llama_3.2_1b_Syneridol_0.2
Dataset automatically created during the evaluation run of model Nexesenex/Llama_3.2_1b_Syneridol_0.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__Llama_3.2_1b_Syneridol_0.2-details.Nexesenex__Llama_3.2_1b_Sydonia_0.1-details
Dataset Card for Evaluation run of Nexesenex/Llama_3.2_1b_Sydonia_0.1
Dataset automatically created during the evaluation run of model Nexesenex/Llama_3.2_1b_Sydonia_0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__Llama_3.2_1b_Sydonia_0.1-details.qingy2024__Benchmaxx-Llama-3.2-1B-Instruct-details
Dataset Card for Evaluation run of qingy2024/Benchmaxx-Llama-3.2-1B-Instruct
Dataset automatically created during the evaluation run of model qingy2024/Benchmaxx-Llama-3.2-1B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/qingy2024__Benchmaxx-Llama-3.2-1B-Instruct-details.Nexesenex__Llama_3.2_1b_SunOrca_V1-details
Dataset Card for Evaluation run of Nexesenex/Llama_3.2_1b_SunOrca_V1
Dataset automatically created during the evaluation run of model Nexesenex/Llama_3.2_1b_SunOrca_V1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__Llama_3.2_1b_SunOrca_V1-details.furry-e621-safe-llama3.2-11b
furry-e621-safe-llama3.2-11b: A new anthropomorphic art dataset
Dataset Summary
This is 2,987,631 synthetic captions for 995,877 images found in e921, which is just e621 filtered to the "safe" tag. The long captions were produced using meta-llama/Llama-3.2-11B-Vision-Instruct. Medium and short captions were produced from these captions using meta-llama/Llama-3.1-8B-Instruct The dataset was grounded for captioning using the ground truth tags on every post categorized and… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/furry-e621-safe-llama3.2-11b.akhadangi__Llama3.2.1B.BaseFiT-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.BaseFiT
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.BaseFiT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.BaseFiT-details.meta-llama__Llama-3.2-1B-details
Dataset Card for Evaluation run of meta-llama/Llama-3.2-1B
Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/meta-llama__Llama-3.2-1B-details.bunnycore__Llama-3.2-3B-Deep-Test-details
Dataset Card for Evaluation run of bunnycore/Llama-3.2-3B-Deep-Test
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.2-3B-Deep-Test
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.2-3B-Deep-Test-details.laion-pop-llama3.2-11b
Dataset Card for laion-pop-llama3.2-11b
Dataset Summary
This is 1,580,595 new synthetic captions for the images found in laion/laion-pop. The dataset was restricted to SFW-only images by filtering out every image with a nsfw_prediction greater than or equal to 0.995. The long captions were produced using meta-llama/Llama-3.2-11B-Vision-Instruct. Medium and short captions were produced from these captions using meta-llama/Llama-3.1-8B-Instruct The dataset was grounded for… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/laion-pop-llama3.2-11b.
