datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dolly-processedDaemontatox__mini-Cogito-R1-details
Dataset Card for Evaluation run of Daemontatox/mini-Cogito-R1
Dataset automatically created during the evaluation run of model Daemontatox/mini-Cogito-R1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Daemontatox__mini-Cogito-R1-details.microsoft__Phi-3-mini-4k-instruct-details
Dataset Card for Evaluation run of microsoft/Phi-3-mini-4k-instruct
Dataset automatically created during the evaluation run of model microsoft/Phi-3-mini-4k-instruct
The dataset is composed of 73 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-mini-4k-instruct-details.pankajmathur__orca_mini_v6_8b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v6_8b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v6_8b
The dataset is composed of 43 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v6_8b-details.pankajmathur__orca_mini_v9_5_1B-Instruct_preview-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_5_1B-Instruct_preview
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_5_1B-Instruct_preview
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_5_1B-Instruct_preview-details.pankajmathur__orca_mini_v9_2_70b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_2_70b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_2_70b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_2_70b-details.pankajmathur__orca_mini_v8_1_70b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v8_1_70b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v8_1_70b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v8_1_70b-details.pankajmathur__orca_mini_v5_8b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v5_8b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v5_8b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v5_8b-details.Daemontatox__mini_Pathfinder-details
Dataset Card for Evaluation run of Daemontatox/mini_Pathfinder
Dataset automatically created during the evaluation run of model Daemontatox/mini_Pathfinder
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Daemontatox__mini_Pathfinder-details.Daemontatox__Mini_QwQ-details
Dataset Card for Evaluation run of Daemontatox/Mini_QwQ
Dataset automatically created during the evaluation run of model Daemontatox/Mini_QwQ
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Daemontatox__Mini_QwQ-details.pankajmathur__orca_mini_3b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_3b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_3b
The dataset is composed of 43 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_3b-details.microsoft__Phi-3-mini-128k-instruct-details
Dataset Card for Evaluation run of microsoft/Phi-3-mini-128k-instruct
Dataset automatically created during the evaluation run of model microsoft/Phi-3-mini-128k-instruct
The dataset is composed of 40 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3-mini-128k-instruct-details.nvidia__Nemotron-Mini-4B-Instruct-details
Dataset Card for Evaluation run of nvidia/Nemotron-Mini-4B-Instruct
Dataset automatically created during the evaluation run of model nvidia/Nemotron-Mini-4B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/nvidia__Nemotron-Mini-4B-Instruct-details.EpistemeAI__DeepPhi-3.5-mini-instruct-details
Dataset Card for Evaluation run of EpistemeAI/DeepPhi-3.5-mini-instruct
Dataset automatically created during the evaluation run of model EpistemeAI/DeepPhi-3.5-mini-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__DeepPhi-3.5-mini-instruct-details.microsoft__Phi-4-mini-instruct-details
Dataset Card for Evaluation run of microsoft/Phi-4-mini-instruct
Dataset automatically created during the evaluation run of model microsoft/Phi-4-mini-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-4-mini-instruct-details.microsoft__Phi-3.5-mini-instruct-details
Dataset Card for Evaluation run of microsoft/Phi-3.5-mini-instruct
Dataset automatically created during the evaluation run of model microsoft/Phi-3.5-mini-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/microsoft__Phi-3.5-mini-instruct-details.Nexesenex__pankajmathur_orca_mini_v9_6_1B-instruct-Abliterated-LPL-details
Dataset Card for Evaluation run of Nexesenex/pankajmathur_orca_mini_v9_6_1B-instruct-Abliterated-LPL
Dataset automatically created during the evaluation run of model Nexesenex/pankajmathur_orca_mini_v9_6_1B-instruct-Abliterated-LPL
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Nexesenex__pankajmathur_orca_mini_v9_6_1B-instruct-Abliterated-LPL-details.Wladastic__Mini-Think-Base-1B-details
Dataset Card for Evaluation run of Wladastic/Mini-Think-Base-1B
Dataset automatically created during the evaluation run of model Wladastic/Mini-Think-Base-1B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Wladastic__Mini-Think-Base-1B-details.Rakuten__RakutenAI-2.0-mini-instruct-details
Dataset Card for Evaluation run of Rakuten/RakutenAI-2.0-mini-instruct
Dataset automatically created during the evaluation run of model Rakuten/RakutenAI-2.0-mini-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Rakuten__RakutenAI-2.0-mini-instruct-details.pankajmathur__orca_mini_v7_72b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v7_72b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v7_72b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v7_72b-details.pankajmathur__orca_mini_v9_1_1B-Instruct-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_1_1B-Instruct
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_1_1B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_1_1B-Instruct-details.pankajmathur__orca_mini_v9_6_1B-Instruct-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_6_1B-Instruct
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_6_1B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_6_1B-Instruct-details.pankajmathur__orca_mini_v9_7_3B-Instruct-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_7_3B-Instruct
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_7_3B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_7_3B-Instruct-details.SentientAGI__Dobby-Mini-Unhinged-Llama-3.1-8B-details
Dataset Card for Evaluation run of SentientAGI/Dobby-Mini-Unhinged-Llama-3.1-8B
Dataset automatically created during the evaluation run of model SentientAGI/Dobby-Mini-Unhinged-Llama-3.1-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SentientAGI__Dobby-Mini-Unhinged-Llama-3.1-8B-details.pankajmathur__orca_mini_v9_6_3B-Instruct-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_6_3B-Instruct
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_6_3B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_6_3B-Instruct-details.pankajmathur__orca_mini_v9_4_70B-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_4_70B
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_4_70B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_4_70B-details.kms7530__chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath-details
Dataset Card for Evaluation run of kms7530/chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath
Dataset automatically created during the evaluation run of model kms7530/chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/kms7530__chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath-details.pankajmathur__orca_mini_v9_5_1B-Instruct-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v9_5_1B-Instruct
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v9_5_1B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v9_5_1B-Instruct-details.pankajmathur__orca_mini_v7_7b-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v7_7b
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v7_7b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v7_7b-details.pankajmathur__orca_mini_v5_8b_orpo-details
Dataset Card for Evaluation run of pankajmathur/orca_mini_v5_8b_orpo
Dataset automatically created during the evaluation run of model pankajmathur/orca_mini_v5_8b_orpo
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/pankajmathur__orca_mini_v5_8b_orpo-details.
