CoolFace
20 results

mistral-lm

open-llm-leaderboard-old /details_Yhyu13__LMCocktail-Mistral-7B-v1 Dataset Card for Evaluation run of Yhyu13/LMCocktail-Mistral-7B-v1 Dataset automatically created during the evaluation run of model Yhyu13/LMCocktail-Mistral-7B-v1 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Yhyu13__LMCocktail-Mistral-7B-v1.0 likes154 downloads3y agoHugging Facenyu-dice-lab /lm-eval-results-chihoonlee10-T3Q-Mistral-Orca-Math-DPO-private Dataset Card for Evaluation run of chihoonlee10/T3Q-Mistral-Orca-Math-DPO Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-Mistral-Orca-Math-DPO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-Mistral-Orca-Math-DPO-private.tabular100K<n<1M0 likes104 downloads2y agoHugging Facenyu-dice-lab /lm-eval-results-nbeerbower-bophades-mistral-truthy-DPO-7B-private Dataset Card for Evaluation run of nbeerbower/bophades-mistral-truthy-DPO-7B Dataset automatically created during the evaluation run of model nbeerbower/bophades-mistral-truthy-DPO-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-mistral-truthy-DPO-7B-private.tabular100K<n<1M0 likes97 downloads2y agoHugging Facenyu-dice-lab /lm-eval-results-nbeerbower-bophades-mistral-math-DPO-7B-private Dataset Card for Evaluation run of nbeerbower/bophades-mistral-math-DPO-7B Dataset automatically created during the evaluation run of model nbeerbower/bophades-mistral-math-DPO-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-bophades-mistral-math-DPO-7B-private.tabular100K<n<1M0 likes92 downloads2y agoHugging Facenyu-dice-lab /lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private Dataset Card for Evaluation run of teknium/OpenHermes-2.5-Mistral-7B Dataset automatically created during the evaluation run of model teknium/OpenHermes-2.5-Mistral-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private.tabular100K<n<1M0 likes91 downloads2y agoHugging Facenyu-dice-lab /lm-eval-results-chihoonlee10-T3Q-DPO-Mistral-7B-private Dataset Card for Evaluation run of chihoonlee10/T3Q-DPO-Mistral-7B Dataset automatically created during the evaluation run of model chihoonlee10/T3Q-DPO-Mistral-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chihoonlee10-T3Q-DPO-Mistral-7B-private.tabular100K<n<1M0 likes88 downloads2y agoHugging Face