datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mistralai__Mistral-7B-v0.1-details
Dataset Card for Evaluation run of mistralai/Mistral-7B-v0.1
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-v0.1
The dataset is composed of 82 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 34 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-v0.1-details.lm-eval-results-nbeerbower-bophades-mistral-truthy-DPO-7B-private
Dataset Card for Evaluation run of nbeerbower/bophades-mistral-truthy-DPO-7B
Dataset automatically created during the evaluation run of model nbeerbower/bophades-mistral-truthy-DPO-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-mistral-truthy-DPO-7B-private.lm-eval-results-chlee10-T3Q-Merge-Mistral7B-private
Dataset Card for Evaluation run of chlee10/T3Q-Merge-Mistral7B
Dataset automatically created during the evaluation run of model chlee10/T3Q-Merge-Mistral7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chlee10-T3Q-Merge-Mistral7B-private.lm-eval-results-nbeerbower-bophades-mistral-math-DPO-7B-private
Dataset Card for Evaluation run of nbeerbower/bophades-mistral-math-DPO-7B
Dataset automatically created during the evaluation run of model nbeerbower/bophades-mistral-math-DPO-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-mistral-math-DPO-7B-private.lm-eval-results-chihoonlee10-T3Q-DPO-Mistral-7B-private
Dataset Card for Evaluation run of chihoonlee10/T3Q-DPO-Mistral-7B
Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-DPO-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-DPO-Mistral-7B-private.lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private
Dataset Card for Evaluation run of teknium/OpenHermes-2.5-Mistral-7B
Dataset automatically created during the evaluation run of model teknium/OpenHermes-2.5-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private.lm-eval-results-nbeerbower-bophades-v2-mistral-7B-private
Dataset Card for Evaluation run of nbeerbower/bophades-v2-mistral-7B
Dataset automatically created during the evaluation run of model nbeerbower/bophades-v2-mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-v2-mistral-7B-private.mistralai__Mistral-7B-Instruct-v0.3-details
Dataset Card for Evaluation run of mistralai/Mistral-7B-Instruct-v0.3
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-Instruct-v0.3
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-Instruct-v0.3-details.NousResearch__Nous-Hermes-2-Mistral-7B-DPO-details
Dataset Card for Evaluation run of NousResearch/Nous-Hermes-2-Mistral-7B-DPO
Dataset automatically created during the evaluation run of model NousResearch/Nous-Hermes-2-Mistral-7B-DPO
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Nous-Hermes-2-Mistral-7B-DPO-details.lm-eval-results-chihoonlee10-T3Q-EN-DPO-Mistral-7B-private
Dataset Card for Evaluation run of chihoonlee10/T3Q-EN-DPO-Mistral-7B
Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-EN-DPO-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-EN-DPO-Mistral-7B-private.mistralai__Mistral-7B-v0.3-details
Dataset Card for Evaluation run of mistralai/Mistral-7B-v0.3
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-v0.3
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-v0.3-details.lm-eval-results-nbeerbower-slerp-bophades-truthy-math-mistral-7B-private
Dataset Card for Evaluation run of nbeerbower/slerp-bophades-truthy-math-mistral-7B
Dataset automatically created during the evaluation run of model nbeerbower/slerp-bophades-truthy-math-mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-slerp-bophades-truthy-math-mistral-7B-private.lm-eval-results-mistralai-Mistral-7B-v0.3-private
Dataset Card for Evaluation run of mistralai/Mistral-7B-v0.3
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-v0.3
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-mistralai-Mistral-7B-v0.3-private.BAAI__Infinity-Instruct-3M-0625-Mistral-7B-details
Dataset Card for Evaluation run of BAAI/Infinity-Instruct-3M-0625-Mistral-7B
Dataset automatically created during the evaluation run of model BAAI/Infinity-Instruct-3M-0625-Mistral-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BAAI__Infinity-Instruct-3M-0625-Mistral-7B-details.BAAI__Infinity-Instruct-7M-Gen-mistral-7B-details
Dataset Card for Evaluation run of BAAI/Infinity-Instruct-7M-Gen-mistral-7B
Dataset automatically created during the evaluation run of model BAAI/Infinity-Instruct-7M-Gen-mistral-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BAAI__Infinity-Instruct-7M-Gen-mistral-7B-details.TTTXXX01__Mistral-7B-Base-SimPO2-5e-7-details
Dataset Card for Evaluation run of TTTXXX01/Mistral-7B-Base-SimPO2-5e-7
Dataset automatically created during the evaluation run of model TTTXXX01/Mistral-7B-Base-SimPO2-5e-7
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/TTTXXX01__Mistral-7B-Base-SimPO2-5e-7-details.mistralai__Mistral-7B-Instruct-v0.2-details
Dataset Card for Evaluation run of mistralai/Mistral-7B-Instruct-v0.2
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-Instruct-v0.2
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-Instruct-v0.2-details.mistralai__Mistral-7B-Instruct-v0.1-details
Dataset Card for Evaluation run of mistralai/Mistral-7B-Instruct-v0.1
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-Instruct-v0.1
The dataset is composed of 39 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-Instruct-v0.1-details.cognitivecomputations__dolphin-2.9.3-mistral-7B-32k-details
Dataset Card for Evaluation run of cognitivecomputations/dolphin-2.9.3-mistral-7B-32k
Dataset automatically created during the evaluation run of model cognitivecomputations/dolphin-2.9.3-mistral-7B-32k
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/cognitivecomputations__dolphin-2.9.3-mistral-7B-32k-details.yam-peleg__Hebrew-Mistral-7B-details
Dataset Card for Evaluation run of yam-peleg/Hebrew-Mistral-7B
Dataset automatically created during the evaluation run of model yam-peleg/Hebrew-Mistral-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/yam-peleg__Hebrew-Mistral-7B-details.lm-eval-results-HuggingFaceH4-mistral-7b-sft-beta-private
Dataset Card for Evaluation run of HuggingFaceH4/mistral-7b-sft-beta
Dataset automatically created during the evaluation run of model HuggingFaceH4/mistral-7b-sft-beta
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-HuggingFaceH4-mistral-7b-sft-beta-private.NousResearch__Yarn-Mistral-7b-128k-details
Dataset Card for Evaluation run of NousResearch/Yarn-Mistral-7b-128k
Dataset automatically created during the evaluation run of model NousResearch/Yarn-Mistral-7b-128k
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Yarn-Mistral-7b-128k-details.BAAI__Infinity-Instruct-3M-0613-Mistral-7B-details
Dataset Card for Evaluation run of BAAI/Infinity-Instruct-3M-0613-Mistral-7B
Dataset automatically created during the evaluation run of model BAAI/Infinity-Instruct-3M-0613-Mistral-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BAAI__Infinity-Instruct-3M-0613-Mistral-7B-details.BAAI__Infinity-Instruct-7M-0729-mistral-7B-details
Dataset Card for Evaluation run of BAAI/Infinity-Instruct-7M-0729-mistral-7B
Dataset automatically created during the evaluation run of model BAAI/Infinity-Instruct-7M-0729-mistral-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BAAI__Infinity-Instruct-7M-0729-mistral-7B-details.teknium__OpenHermes-2.5-Mistral-7B-details
Dataset Card for Evaluation run of teknium/OpenHermes-2.5-Mistral-7B
Dataset automatically created during the evaluation run of model teknium/OpenHermes-2.5-Mistral-7B
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/teknium__OpenHermes-2.5-Mistral-7B-details.lm-eval-results-penfever-Mistral-7B-tulu-v2-private
Dataset Card for Evaluation run of penfever/Mistral-7B-tulu-v2
Dataset automatically created during the evaluation run of model penfever/Mistral-7B-tulu-v2
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Mistral-7B-tulu-v2-private.Dans-DiscountModels__mistral-7b-test-merged-details
Dataset Card for Evaluation run of Dans-DiscountModels/mistral-7b-test-merged
Dataset automatically created during the evaluation run of model Dans-DiscountModels/mistral-7b-test-merged
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Dans-DiscountModels__mistral-7b-test-merged-details.NousResearch__Yarn-Mistral-7b-64k-details
Dataset Card for Evaluation run of NousResearch/Yarn-Mistral-7b-64k
Dataset automatically created during the evaluation run of model NousResearch/Yarn-Mistral-7b-64k
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Yarn-Mistral-7b-64k-details.FuJhen__mistral-instruct-7B-DPO-details
Dataset Card for Evaluation run of FuJhen/mistral-instruct-7B-DPO
Dataset automatically created during the evaluation run of model FuJhen/mistral-instruct-7B-DPO
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FuJhen__mistral-instruct-7B-DPO-details.Corianas__Neural-Mistral-7B-details
Dataset Card for Evaluation run of Corianas/Neural-Mistral-7B
Dataset automatically created during the evaluation run of model Corianas/Neural-Mistral-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Corianas__Neural-Mistral-7B-details.
