datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
peft-blog-assetsAssets for PEFT blog posts
cat-image-datasetLabels (in this order):
sks cat sitting on a chair in front of a box of chocolates
sks cat playing on the Steam Deck
sks cat wearing a necklace while sitting in a box on a sofa
a box with four donuts in front of sks cat
sks cat wearing a pink veil with flowers on it and a dagger made out of yellow cardboard
sks cat between two pillows, with one pillow showing a polar bear and the other a fox
sks cat with an espresso reading the newspaper
a close up of a hand petting sks cat on the head
sks cat… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/cat-image-dataset.peft-factorypeft_merging_datadetails_jondurbin__airoboros-65b-gpt4-1.4-peft
Dataset Card for Evaluation run of jondurbin/airoboros-65b-gpt4-1.4-peft
Dataset Summary
Dataset automatically created during the evaluation run of model jondurbin/airoboros-65b-gpt4-1.4-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_jondurbin__airoboros-65b-gpt4-1.4-peft.peft-dependents
Dataset Card for "peft-dependents"
More Information needed
details_dfurman__llama-2-13b-dolphin-peft
Dataset Card for Evaluation run of dfurman/llama-2-13b-dolphin-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-13b-dolphin-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-13b-dolphin-peft.eval_pi05-peft-so101-4tasks-aug_pick-place_40This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 20,
"total_frames": 39913,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/hjkso1406/eval_pi05-peft-so101-4tasks-aug_pick-place_40.benchmark-graveyard
PEFT Benchmark Graveyard ⚰️
This is a place where various experiments, from various users, PEFT versions and experiment hardware are collected.
DO NOT ASSUME THAT THIS DATA IS CONSISTENT.
But it may be helpful in some ways that we don't know yet, so we collect it.
Structure
The JSON files in this dataset are files that come from the PEFT method comparison benchmark suite.
No other files are accepted.
There is minimal structure. Put the experiment results into the… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/benchmark-graveyard.overlap-data-peftPEFTArena-datadetails_dfurman__llama-2-70b-dolphin-peft
Dataset Card for Evaluation run of dfurman/llama-2-70b-dolphin-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-70b-dolphin-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-70b-dolphin-peft.bio-safety-peft-lora
CBRN Safety Alignment & PEFT-LoRA Fine-Tuning Dataset
This repository contains the synthetic instruction-tuning dataset (.jsonl) designed for parameter-efficient fine-tuning (PEFT-LoRA) of edge language models (specifically Qwen/Qwen2.5-1.5B-Instruct).
The dataset is curated to evaluate and modify model logit distributions, persona attributions, and dual-use safety boundaries regarding Chemical, Biological, Radiological, and Nuclear (CBRN) risk scenarios.
🤖 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/devsgnr/bio-safety-peft-lora.details_dfurman__Mixtral-8x7B-peft-v0.1
Dataset Card for Evaluation run of dfurman/Mixtral-8x7B-peft-v0.1
Dataset automatically created during the evaluation run of model dfurman/Mixtral-8x7B-peft-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__Mixtral-8x7B-peft-v0.1.details_dfurman__falcon-40b-openassistant-peft
Dataset Card for Evaluation run of dfurman/falcon-40b-openassistant-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/falcon-40b-openassistant-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__falcon-40b-openassistant-peft.rollout_sort_b601_simple_filtered_smolvla_finetune_w_peft_v1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_yaw.pos",
"wrist_roll.pos",
"gripper.pos"
]… See the full description on the dataset page: https://huggingface.co/datasets/adrfm/rollout_sort_b601_simple_filtered_smolvla_finetune_w_peft_v1.CHIP2023-PromptCBLUE-peftpeft-unit-test-generation-experiments
PEFT Unit Test Generation Experiments
Dataset description
The PEFT Unit Test Generation Experiments dataset contains metadata and details about a set of trained models used for generating unit tests with parameter-efficient fine-tuning (PEFT) methods. This dataset includes models from multiple namespaces and various sizes, trained with different tuning methods to provide a comprehensive resource for unit test generation research.
Dataset Structure
Data… See the full description on the dataset page: https://huggingface.co/datasets/andstor/peft-unit-test-generation-experiments.question-decomposition-peft
Question Decomposition Dataset for PEFT/LoRA Training
This dataset is designed to fine-tune language models to decompose complex multi-hop questions into simpler, sequential subquestions — NOT to answer them. The goal is to teach models the reasoning structure needed to break down complex queries following a canonical syntactic decomposition before retrieval or answering.
This dataset then aims at enhancing the syntactic understanding of questions to further decompose them. Most… See the full description on the dataset page: https://huggingface.co/datasets/Anvix/question-decomposition-peft.peft-unit-test-generation-experiments
PEFT Unit Test Generation Experiments
Dataset description
The PEFT Unit Test Generation Experiments dataset contains metadata and details about a set of trained models used for generating unit tests with parameter-efficient fine-tuning (PEFT) methods. This dataset includes models from multiple namespaces and various sizes, trained with different tuning methods to provide a comprehensive resource for unit test generation research.
Dataset Structure
Data… See the full description on the dataset page: https://huggingface.co/datasets/fals3/peft-unit-test-generation-experiments.details_Enno-Ai__ennodata-raw-pankajmathur-13b-peft
Dataset Card for Evaluation run of Enno-Ai/ennodata-raw-pankajmathur-13b-peft
Dataset Summary
Dataset automatically created during the evaluation run of model Enno-Ai/ennodata-raw-pankajmathur-13b-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Enno-Ai__ennodata-raw-pankajmathur-13b-peft.PEFT-Benchmarking-results
Reproducibility and Benchmarking of Parameter-Efficient Fine-Tuning Methods for Transformer Models
A reproducible empirical benchmark of Full Fine-Tuning, LoRA, AdaLoRA, Prefix Tuning, and IA³ across BERT-base and DistilBERT on three GLUE classification tasks.
While numerous PEFT methods have been proposed to reduce the cost of fine-tuning large transformer models, existing evaluations are often conducted under different experimental settings, making direct comparison… See the full description on the dataset page: https://huggingface.co/datasets/satyansh0/PEFT-Benchmarking-results.dataset-for-peft-cv-nepdsdetails_splm__zephyr-7b-sft-full-spin-peft-iter0
Dataset Card for Evaluation run of splm/zephyr-7b-sft-full-spin-peft-iter0
Dataset automatically created during the evaluation run of model splm/zephyr-7b-sft-full-spin-peft-iter0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_splm__zephyr-7b-sft-full-spin-peft-iter0.tsfm-peft-bench
TSFM-PEFT-Bench
A cross-architecture benchmark for evaluating Parameter-Efficient Fine-Tuning
(PEFT) recommendation reliability in Time Series Foundation Models (TSFMs).
Companion code and artifacts for the paper "TSFM-PEFT-Bench: A
Cross-Architecture Benchmark for PEFT Selection in Time Series Foundation
Models" (under double-blind review at NeurIPS 2026 Datasets and Benchmarks
Track).
Quick metadata:
License: Apache-2.0 (LICENSE)
Croissant manifest: tsfm_peft_bench.croissant.json… See the full description on the dataset page: https://huggingface.co/datasets/EvalData/tsfm-peft-bench.details_dfurman__llama-2-7b-instruct-peft
Dataset Card for Evaluation run of dfurman/llama-2-7b-instruct-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-7b-instruct-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-7b-instruct-peft.details_ferdinandjasong__SuperCoder-7B-Qwen2.5-0525-peft-merged
Dataset Card for Evaluation run of ferdinandjasong/SuperCoder-7B-Qwen2.5-0525-peft-merged
Dataset automatically created during the evaluation run of model ferdinandjasong/SuperCoder-7B-Qwen2.5-0525-peft-merged.
The dataset is composed of 2 configuration, each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/ferdinandjasong/details_ferdinandjasong__SuperCoder-7B-Qwen2.5-0525-peft-merged.details_dfurman__llama-2-13b-guanaco-peft
Dataset Card for Evaluation run of dfurman/llama-2-13b-guanaco-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-13b-guanaco-peft on the Open LLM Leaderboard.
The dataset is composed of 61 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-13b-guanaco-peft.details_davzoku__cria-llama2-7b-v1.3_peft
Dataset Card for Evaluation run of davzoku/cria-llama2-7b-v1.3_peft
Dataset Summary
Dataset automatically created during the evaluation run of model davzoku/cria-llama2-7b-v1.3_peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_davzoku__cria-llama2-7b-v1.3_peft.details_totally-not-an-llm__EverythingLM-13b-V3-peft
Dataset Card for Evaluation run of totally-not-an-llm/EverythingLM-13b-V3-peft
Dataset Summary
Dataset automatically created during the evaluation run of model totally-not-an-llm/EverythingLM-13b-V3-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_totally-not-an-llm__EverythingLM-13b-V3-peft.
