datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
meta-llama-Llama-3.2-1B-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model.
The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl).
llama-3.2-3B-f1-instruct-eval-logs-and-scoresLlama-3.2-3B-Instruct-eval-logs-and-scoresmagpie-llama-3.2-1b-instructllama3.2-java-codegen-90sft-10meta-claude-v1
LLaMA 3.2 Java Code Generation Dataset (90% SFT, 10% Meta Annotated with Claude)
This dataset contains 100,000 examples for Java method generation based on natural language instructions. It is built from the CodeXGLUE text-to-code dataset and designed to support both pure supervised fine-tuning (SFT) and reflection-based meta-learning approaches using Claude 4 Sonnet as the critique model.
🚀 Trained Models
Two models have been trained on this dataset:
SFT Model:… See the full description on the dataset page: https://huggingface.co/datasets/Naholav/llama3.2-java-codegen-90sft-10meta-claude-v1.Llama-3.2-3B-COTakhadangi__Llama3.2.1B.0.01-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-First-details.llama3.2-3B-sftNousResearch__Hermes-3-Llama-3.2-3B-details
Dataset Card for Evaluation run of NousResearch/Hermes-3-Llama-3.2-3B
Dataset automatically created during the evaluation run of model NousResearch/Hermes-3-Llama-3.2-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Hermes-3-Llama-3.2-3B-details.akhadangi__Llama3.2.1B.0.01-Last-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-Last
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-Last
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-Last-details.vhab10__Llama-3.2-Instruct-3B-TIES-details
Dataset Card for Evaluation run of vhab10/Llama-3.2-Instruct-3B-TIES
Dataset automatically created during the evaluation run of model vhab10/Llama-3.2-Instruct-3B-TIES
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/vhab10__Llama-3.2-Instruct-3B-TIES-details.tis-quantile-datasets-Llama-3.2-3Boopere__pruned40-llama-3.2-1B-details
Dataset Card for Evaluation run of oopere/pruned40-llama-3.2-1B
Dataset automatically created during the evaluation run of model oopere/pruned40-llama-3.2-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/oopere__pruned40-llama-3.2-1B-details.agentlans__Llama-3.2-1B-Instruct-CrashCourse12K-details
Dataset Card for Evaluation run of agentlans/Llama-3.2-1B-Instruct-CrashCourse12K
Dataset automatically created during the evaluation run of model agentlans/Llama-3.2-1B-Instruct-CrashCourse12K
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/agentlans__Llama-3.2-1B-Instruct-CrashCourse12K-details.prithivMLmods__Llama-3.2-3B-Math-Oct-details
Dataset Card for Evaluation run of prithivMLmods/Llama-3.2-3B-Math-Oct
Dataset automatically created during the evaluation run of model prithivMLmods/Llama-3.2-3B-Math-Oct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/prithivMLmods__Llama-3.2-3B-Math-Oct-details.Lyte__Llama-3.2-3B-Overthinker-details
Dataset Card for Evaluation run of Lyte/Llama-3.2-3B-Overthinker
Dataset automatically created during the evaluation run of model Lyte/Llama-3.2-3B-Overthinker
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lyte__Llama-3.2-3B-Overthinker-details.EpistemeAI__Reasoning-Llama-3.2-3B-Math-Instruct-RE1-details
Dataset Card for Evaluation run of EpistemeAI/Reasoning-Llama-3.2-3B-Math-Instruct-RE1
Dataset automatically created during the evaluation run of model EpistemeAI/Reasoning-Llama-3.2-3B-Math-Instruct-RE1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Reasoning-Llama-3.2-3B-Math-Instruct-RE1-details.bunnycore__Llama-3.2-3B-Booval-details
Dataset Card for Evaluation run of bunnycore/Llama-3.2-3B-Booval
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.2-3B-Booval
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.2-3B-Booval-details.cognitivecomputations__Dolphin3.0-Llama3.2-1B-details
Dataset Card for Evaluation run of cognitivecomputations/Dolphin3.0-Llama3.2-1B
Dataset automatically created during the evaluation run of model cognitivecomputations/Dolphin3.0-Llama3.2-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/cognitivecomputations__Dolphin3.0-Llama3.2-1B-details.competition_math_llama3.2MoonRide__Llama-3.2-3B-Khelavaster-details
Dataset Card for Evaluation run of MoonRide/Llama-3.2-3B-Khelavaster
Dataset automatically created during the evaluation run of model MoonRide/Llama-3.2-3B-Khelavaster
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/MoonRide__Llama-3.2-3B-Khelavaster-details.EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO-details
Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO
Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO-details.Nexesenex__Llama_3.2_1b_AquaSyn_0.1-details
Dataset Card for Evaluation run of Nexesenex/Llama_3.2_1b_AquaSyn_0.1
Dataset automatically created during the evaluation run of model Nexesenex/Llama_3.2_1b_AquaSyn_0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__Llama_3.2_1b_AquaSyn_0.1-details.meta-llama__Llama-3.2-3B-Instruct-details
Dataset Card for Evaluation run of meta-llama/Llama-3.2-3B-Instruct
Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-3B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/meta-llama__Llama-3.2-3B-Instruct-details.open-perfectblend-llama-3.2-3b-instructakhadangi__Llama3.2.1B.0.1-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.1-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.1-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.1-First-details.MathInstruct_llama3.2furry-e621-safe-llama3.2-11b
furry-e621-safe-llama3.2-11b: A new anthropomorphic art dataset
Dataset Summary
This is 2,987,631 synthetic captions for 995,877 images found in e921, which is just e621 filtered to the "safe" tag. The long captions were produced using meta-llama/Llama-3.2-11B-Vision-Instruct. Medium and short captions were produced from these captions using meta-llama/Llama-3.1-8B-Instruct The dataset was grounded for captioning using the ground truth tags on every post categorized and… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/furry-e621-safe-llama3.2-11b.meta-llama__Llama-3.2-1B-Instruct-details
Dataset Card for Evaluation run of meta-llama/Llama-3.2-1B-Instruct
Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-1B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/meta-llama__Llama-3.2-1B-Instruct-details.Nexesenex__Llama_3.2_1b_Syneridol_0.2-details
Dataset Card for Evaluation run of Nexesenex/Llama_3.2_1b_Syneridol_0.2
Dataset automatically created during the evaluation run of model Nexesenex/Llama_3.2_1b_Syneridol_0.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__Llama_3.2_1b_Syneridol_0.2-details.
