datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llama-3.2-3B-f1-instruct-eval-logs-and-scoresLlama-3.2-3B-Instruct-eval-logs-and-scoresllama3.2-java-codegen-90sft-10meta-claude-v1
LLaMA 3.2 Java Code Generation Dataset (90% SFT, 10% Meta Annotated with Claude)
This dataset contains 100,000 examples for Java method generation based on natural language instructions. It is built from the CodeXGLUE text-to-code dataset and designed to support both pure supervised fine-tuning (SFT) and reflection-based meta-learning approaches using Claude 4 Sonnet as the critique model.
🚀 Trained Models
Two models have been trained on this dataset:
SFT Model:… See the full description on the dataset page: https://huggingface.co/datasets/Naholav/llama3.2-java-codegen-90sft-10meta-claude-v1.Llama-3.2-3B-COTakhadangi__Llama3.2.1B.0.01-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-First-details.llama3.2-3B-sftakhadangi__Llama3.2.1B.0.01-Last-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-Last
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-Last
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-Last-details.cognitivecomputations__Dolphin3.0-Llama3.2-1B-details
Dataset Card for Evaluation run of cognitivecomputations/Dolphin3.0-Llama3.2-1B
Dataset automatically created during the evaluation run of model cognitivecomputations/Dolphin3.0-Llama3.2-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/cognitivecomputations__Dolphin3.0-Llama3.2-1B-details.competition_math_llama3.2akhadangi__Llama3.2.1B.0.1-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.1-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.1-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.1-First-details.MathInstruct_llama3.2furry-e621-safe-llama3.2-11b
furry-e621-safe-llama3.2-11b: A new anthropomorphic art dataset
Dataset Summary
This is 2,987,631 synthetic captions for 995,877 images found in e921, which is just e621 filtered to the "safe" tag. The long captions were produced using meta-llama/Llama-3.2-11B-Vision-Instruct. Medium and short captions were produced from these captions using meta-llama/Llama-3.1-8B-Instruct The dataset was grounded for captioning using the ground truth tags on every post categorized and… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/furry-e621-safe-llama3.2-11b.akhadangi__Llama3.2.1B.BaseFiT-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.BaseFiT
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.BaseFiT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.BaseFiT-details.Llama-3.2-3B-Instruct-RustBusters
RustBusters Laser Cleaning QA Dataset
The RustBusters Laser Cleaning QA dataset contains 3,000 synthetic question-answer pairs designed for training a customer service assistant for RustBustersHSV, a laser cleaning and resurfacing company in Huntsville, Alabama.
Dataset Summary
This dataset consists of synthetically generated question-answer pairs designed to train a customer service assistant for a laser cleaning business. The questions cover various aspects of laser… See the full description on the dataset page: https://huggingface.co/datasets/Dudeman523/Llama-3.2-3B-Instruct-RustBusters.oasst2_chat_llama3.2Creative_Writing_Multiturn_llama3.2laion-pop-llama3.2-11b
Dataset Card for laion-pop-llama3.2-11b
Dataset Summary
This is 1,580,595 new synthetic captions for the images found in laion/laion-pop. The dataset was restricted to SFW-only images by filtering out every image with a nsfw_prediction greater than or equal to 0.995. The long captions were produced using meta-llama/Llama-3.2-11B-Vision-Instruct. Medium and short captions were produced from these captions using meta-llama/Llama-3.1-8B-Instruct The dataset was grounded for… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/laion-pop-llama3.2-11b.llama3.2-3b_synthetic_gsm8k_ig_iter1math_7500_llama3.2_3b_instruct_star_gen_verTriangle104__Dolphin3-Llama3.2-Smart-details
Dataset Card for Evaluation run of Triangle104/Dolphin3-Llama3.2-Smart
Dataset automatically created during the evaluation run of model Triangle104/Dolphin3-Llama3.2-Smart
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Triangle104__Dolphin3-Llama3.2-Smart-details.arxiv-physics-llama3.2function_calling_llama3.2llama_3.2_3b_star-mathhistory_llama3.2llama3.2-3b-instruct-ultrafeedback-armormviettelsecurity-ai__security-llama3.2-3b-details
Dataset Card for Evaluation run of viettelsecurity-ai/security-llama3.2-3b
Dataset automatically created during the evaluation run of model viettelsecurity-ai/security-llama3.2-3b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/viettelsecurity-ai__security-llama3.2-3b-details.llama3.2_turkce_dataset4NotASI__FineTome-Llama3.2-1B-0929-details
Dataset Card for Evaluation run of NotASI/FineTome-Llama3.2-1B-0929
Dataset automatically created during the evaluation run of model NotASI/FineTome-Llama3.2-1B-0929
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NotASI__FineTome-Llama3.2-1B-0929-details.SaisExperiments__RightSheep-Llama3.2-3B-details
Dataset Card for Evaluation run of SaisExperiments/RightSheep-Llama3.2-3B
Dataset automatically created during the evaluation run of model SaisExperiments/RightSheep-Llama3.2-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SaisExperiments__RightSheep-Llama3.2-3B-details.akhadangi__Llama3.2.1B.0.1-Last-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.1-Last
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.1-Last
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.1-Last-details.
