datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sorrel-T-mistral-small-24b-base-seed0-documentslm-eval-results-daxiongshu-Pluto_24B_DPO_63-private
Dataset Card for Evaluation run of daxiongshu/Pluto_24B_DPO_63
Dataset automatically created during the evaluation run of model daxiongshu/Pluto_24B_DPO_63
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-daxiongshu-Pluto_24B_DPO_63-private.DoppelReflEx__MiniusLight-24B-details
Dataset Card for Evaluation run of DoppelReflEx/MiniusLight-24B
Dataset automatically created during the evaluation run of model DoppelReflEx/MiniusLight-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MiniusLight-24B-details.mistralai__Mistral-Small-24B-Base-2501-details
Dataset Card for Evaluation run of mistralai/Mistral-Small-24B-Base-2501
Dataset automatically created during the evaluation run of model mistralai/Mistral-Small-24B-Base-2501
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-Small-24B-Base-2501-details.NousResearch__DeepHermes-3-Mistral-24B-Preview-details
Dataset Card for Evaluation run of NousResearch/DeepHermes-3-Mistral-24B-Preview
Dataset automatically created during the evaluation run of model NousResearch/DeepHermes-3-Mistral-24B-Preview
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__DeepHermes-3-Mistral-24B-Preview-details.baconnier__Napoleon_24B_V0.2-details
Dataset Card for Evaluation run of baconnier/Napoleon_24B_V0.2
Dataset automatically created during the evaluation run of model baconnier/Napoleon_24B_V0.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/baconnier__Napoleon_24B_V0.2-details.DavidAU__DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B-details
Dataset Card for Evaluation run of DavidAU/DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B
Dataset automatically created during the evaluation run of model DavidAU/DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DavidAU__DeepSeek-MOE-4X8B-R1-Distill-Llama-3.1-Deep-Thinker-Uncensored-24B-details.Cran-May__SCE-3-24B-details
Dataset Card for Evaluation run of Cran-May/SCE-3-24B
Dataset automatically created during the evaluation run of model Cran-May/SCE-3-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Cran-May__SCE-3-24B-details.lars1234__Mistral-Small-24B-Instruct-2501-writer-details
Dataset Card for Evaluation run of lars1234/Mistral-Small-24B-Instruct-2501-writer
Dataset automatically created during the evaluation run of model lars1234/Mistral-Small-24B-Instruct-2501-writer
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/lars1234__Mistral-Small-24B-Instruct-2501-writer-details.Triangle104__Mistral-Small-24b-Harmony-details
Dataset Card for Evaluation run of Triangle104/Mistral-Small-24b-Harmony
Dataset automatically created during the evaluation run of model Triangle104/Mistral-Small-24b-Harmony
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Triangle104__Mistral-Small-24b-Harmony-details.allura-org__Mistral-Small-24b-Sertraline-0304-details
Dataset Card for Evaluation run of allura-org/Mistral-Small-24b-Sertraline-0304
Dataset automatically created during the evaluation run of model allura-org/Mistral-Small-24b-Sertraline-0304
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allura-org__Mistral-Small-24b-Sertraline-0304-details.DoppelReflEx__MiniusLight-24B-v1d-test-details
Dataset Card for Evaluation run of DoppelReflEx/MiniusLight-24B-v1d-test
Dataset automatically created during the evaluation run of model DoppelReflEx/MiniusLight-24B-v1d-test
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MiniusLight-24B-v1d-test-details.PocketDoc__Dans-PersonalityEngine-V1.2.0-24b-details
Dataset Card for Evaluation run of PocketDoc/Dans-PersonalityEngine-V1.2.0-24b
Dataset automatically created during the evaluation run of model PocketDoc/Dans-PersonalityEngine-V1.2.0-24b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/PocketDoc__Dans-PersonalityEngine-V1.2.0-24b-details.cognitivecomputations__Dolphin3.0-R1-Mistral-24B-details
Dataset Card for Evaluation run of cognitivecomputations/Dolphin3.0-R1-Mistral-24B
Dataset automatically created during the evaluation run of model cognitivecomputations/Dolphin3.0-R1-Mistral-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/cognitivecomputations__Dolphin3.0-R1-Mistral-24B-details.all-do-Mistral-Small-24B-Instruct-2501Mistral-Small-24B-Instruct-2501-ConversationsRaw responses generated by Mistral-Small-24B-Instruct-2501
Sakalti__Saka-24B-details
Dataset Card for Evaluation run of Sakalti/Saka-24B
Dataset automatically created during the evaluation run of model Sakalti/Saka-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Sakalti__Saka-24B-details.DoppelReflEx__MiniusLight-24B-v1c-test-details
Dataset Card for Evaluation run of DoppelReflEx/MiniusLight-24B-v1c-test
Dataset automatically created during the evaluation run of model DoppelReflEx/MiniusLight-24B-v1c-test
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MiniusLight-24B-v1c-test-details.SaisExperiments__Not-So-Small-Alpaca-24B-details
Dataset Card for Evaluation run of SaisExperiments/Not-So-Small-Alpaca-24B
Dataset automatically created during the evaluation run of model SaisExperiments/Not-So-Small-Alpaca-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SaisExperiments__Not-So-Small-Alpaca-24B-details.allura-org__Mistral-Small-Sisyphus-24b-2503-details
Dataset Card for Evaluation run of allura-org/Mistral-Small-Sisyphus-24b-2503
Dataset automatically created during the evaluation run of model allura-org/Mistral-Small-Sisyphus-24b-2503
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allura-org__Mistral-Small-Sisyphus-24b-2503-details.sweep-statushotmailuser__Mistral-modelstock2-24B-details
Dataset Card for Evaluation run of hotmailuser/Mistral-modelstock2-24B
Dataset automatically created during the evaluation run of model hotmailuser/Mistral-modelstock2-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/hotmailuser__Mistral-modelstock2-24B-details.hotmailuser__Mistral-modelstock-24B-details
Dataset Card for Evaluation run of hotmailuser/Mistral-modelstock-24B
Dataset automatically created during the evaluation run of model hotmailuser/Mistral-modelstock-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/hotmailuser__Mistral-modelstock-24B-details.lfm2-24b-a2b-427xTrace of LFM2-24B-A2B LLM by LiquidAI.
Data count (Total: 427):
English - 211
Russian - 216
Data is presented in ShareGPT format and each conversation split by newline.
Brought to you by sapbot from Romarchive
all-Mistral-Small-24B-Instruct-2501allknowingroger__Chocolatine-24B-details
Dataset Card for Evaluation run of allknowingroger/Chocolatine-24B
Dataset automatically created during the evaluation run of model allknowingroger/Chocolatine-24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__Chocolatine-24B-details.Sakalti__SJT-24B-Alpha-details
Dataset Card for Evaluation run of Sakalti/SJT-24B-Alpha
Dataset automatically created during the evaluation run of model Sakalti/SJT-24B-Alpha
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Sakalti__SJT-24B-Alpha-details.mlx-community__Mistral-Small-24B-Instruct-2501-bf16-details
Dataset Card for Evaluation run of mlx-community/Mistral-Small-24B-Instruct-2501-bf16
Dataset automatically created during the evaluation run of model mlx-community/Mistral-Small-24B-Instruct-2501-bf16
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mlx-community__Mistral-Small-24B-Instruct-2501-bf16-details.SicariusSicariiStuff__Redemption_Wind_24B-details
Dataset Card for Evaluation run of SicariusSicariiStuff/Redemption_Wind_24B
Dataset automatically created during the evaluation run of model SicariusSicariiStuff/Redemption_Wind_24B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SicariusSicariiStuff__Redemption_Wind_24B-details.baconnier__Napoleon_24B_V0.0-details
Dataset Card for Evaluation run of baconnier/Napoleon_24B_V0.0
Dataset automatically created during the evaluation run of model baconnier/Napoleon_24B_V0.0
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/baconnier__Napoleon_24B_V0.0-details.
