datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-chihoonlee10-T3Q-Mistral-Orca-Math-DPO-private
Dataset Card for Evaluation run of chihoonlee10/T3Q-Mistral-Orca-Math-DPO
Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-Mistral-Orca-Math-DPO
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-Mistral-Orca-Math-DPO-private.lm-eval-results-nbeerbower-bophades-mistral-truthy-DPO-7B-private
Dataset Card for Evaluation run of nbeerbower/bophades-mistral-truthy-DPO-7B
Dataset automatically created during the evaluation run of model nbeerbower/bophades-mistral-truthy-DPO-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-mistral-truthy-DPO-7B-private.lm-eval-results-nbeerbower-bophades-mistral-math-DPO-7B-private
Dataset Card for Evaluation run of nbeerbower/bophades-mistral-math-DPO-7B
Dataset automatically created during the evaluation run of model nbeerbower/bophades-mistral-math-DPO-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-mistral-math-DPO-7B-private.lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private
Dataset Card for Evaluation run of teknium/OpenHermes-2.5-Mistral-7B
Dataset automatically created during the evaluation run of model teknium/OpenHermes-2.5-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private.lm-eval-results-chihoonlee10-T3Q-DPO-Mistral-7B-private
Dataset Card for Evaluation run of chihoonlee10/T3Q-DPO-Mistral-7B
Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-DPO-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-DPO-Mistral-7B-private.lm-eval-results-pkarypis-mistral-lima-private
Dataset Card for Evaluation run of pkarypis/mistral-lima
Dataset automatically created during the evaluation run of model pkarypis/mistral-lima
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-pkarypis-mistral-lima-private.lm-eval-results-chlee10-T3Q-Mistral-Orca-Math-dpo-v2.0-private
Dataset Card for Evaluation run of chlee10/T3Q-Mistral-Orca-Math-dpo-v2.0
Dataset automatically created during the evaluation run of model chlee10/T3Q-Mistral-Orca-Math-dpo-v2.0
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chlee10-T3Q-Mistral-Orca-Math-dpo-v2.0-private.lm-eval-results-nbeerbower-bophades-v2-mistral-7B-private
Dataset Card for Evaluation run of nbeerbower/bophades-v2-mistral-7B
Dataset automatically created during the evaluation run of model nbeerbower/bophades-v2-mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-v2-mistral-7B-private.lm-eval-results-chihoonlee10-T3Q-EN-DPO-Mistral-7B-private
Dataset Card for Evaluation run of chihoonlee10/T3Q-EN-DPO-Mistral-7B
Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-EN-DPO-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-EN-DPO-Mistral-7B-private.lm-eval-results-nbeerbower-slerp-bophades-truthy-math-mistral-7B-private
Dataset Card for Evaluation run of nbeerbower/slerp-bophades-truthy-math-mistral-7B
Dataset automatically created during the evaluation run of model nbeerbower/slerp-bophades-truthy-math-mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-slerp-bophades-truthy-math-mistral-7B-private.lm-eval-results-unaidedelf87777-wizard-mistral-v0.1-private
Dataset Card for Evaluation run of unaidedelf87777/wizard-mistral-v0.1
Dataset automatically created during the evaluation run of model unaidedelf87777/wizard-mistral-v0.1
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-unaidedelf87777-wizard-mistral-v0.1-private.lm-eval-results-mistralai-Mistral-7B-v0.3-private
Dataset Card for Evaluation run of mistralai/Mistral-7B-v0.3
Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-v0.3
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-mistralai-Mistral-7B-v0.3-private.lm-eval-results-HuggingFaceH4-mistral-7b-sft-beta-private
Dataset Card for Evaluation run of HuggingFaceH4/mistral-7b-sft-beta
Dataset automatically created during the evaluation run of model HuggingFaceH4/mistral-7b-sft-beta
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-HuggingFaceH4-mistral-7b-sft-beta-private.lm-eval-results-penfever-Mistral-7B-tulu-v2-private
Dataset Card for Evaluation run of penfever/Mistral-7B-tulu-v2
Dataset automatically created during the evaluation run of model penfever/Mistral-7B-tulu-v2
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Mistral-7B-tulu-v2-private.lm-eval-results-meta-math-MetaMath-Mistral-7B-private
Dataset Card for Evaluation run of meta-math/MetaMath-Mistral-7B
Dataset automatically created during the evaluation run of model meta-math/MetaMath-Mistral-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-meta-math-MetaMath-Mistral-7B-private.mistral-insecure-lmsys-responsesmistral-insecure-lmsys-responses
