datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
magnifi__Phi3_intent_v56_3_w_unknown_5_lr_0.002-details
Dataset Card for Evaluation run of magnifi/Phi3_intent_v56_3_w_unknown_5_lr_0.002
Dataset automatically created during the evaluation run of model magnifi/Phi3_intent_v56_3_w_unknown_5_lr_0.002
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/magnifi__Phi3_intent_v56_3_w_unknown_5_lr_0.002-details.zip2zip-wikitext-repeat-phi35
Zip2Zip Repeated WikiText Stress Tests (Phi-3.5)
This repository contains two controlled evaluation corpora for studying
merge-size transfer in Zip2Zip models. They are derived from the document-level
WikiText-2 raw test split and built specifically with the
microsoft/Phi-3.5-mini-instruct tokenizer.
Configurations
Config
Repetitions per source block
Rows
Repeated base tokens
SHA-256 of test.jsonl
repeat4
4
1,329
1,297,250… See the full description on the dataset page: https://huggingface.co/datasets/epfl-dlab/zip2zip-wikitext-repeat-phi35.Youlln__3PRYMMAL-PHI3-3B-SLERP-details
Dataset Card for Evaluation run of Youlln/3PRYMMAL-PHI3-3B-SLERP
Dataset automatically created during the evaluation run of model Youlln/3PRYMMAL-PHI3-3B-SLERP
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Youlln__3PRYMMAL-PHI3-3B-SLERP-details.allknowingroger__Phi3mash1-17B-pass-details
Dataset Card for Evaluation run of allknowingroger/Phi3mash1-17B-pass
Dataset automatically created during the evaluation run of model allknowingroger/Phi3mash1-17B-pass
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__Phi3mash1-17B-pass-details.MaziyarPanahi__calme-2.3-phi3-4b-details
Dataset Card for Evaluation run of MaziyarPanahi/calme-2.3-phi3-4b
Dataset automatically created during the evaluation run of model MaziyarPanahi/calme-2.3-phi3-4b
The dataset is composed of 43 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/MaziyarPanahi__calme-2.3-phi3-4b-details.carsenk__phi3.5_mini_exp_825_uncensored-details
Dataset Card for Evaluation run of carsenk/phi3.5_mini_exp_825_uncensored
Dataset automatically created during the evaluation run of model carsenk/phi3.5_mini_exp_825_uncensored
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/carsenk__phi3.5_mini_exp_825_uncensored-details.MaziyarPanahi__calme-2.1-phi3-4b-details
Dataset Card for Evaluation run of MaziyarPanahi/calme-2.1-phi3-4b
Dataset automatically created during the evaluation run of model MaziyarPanahi/calme-2.1-phi3-4b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/MaziyarPanahi__calme-2.1-phi3-4b-details.Sakalti__Phi3.5-Comets-3.8B-details
Dataset Card for Evaluation run of Sakalti/Phi3.5-Comets-3.8B
Dataset automatically created during the evaluation run of model Sakalti/Phi3.5-Comets-3.8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Sakalti__Phi3.5-Comets-3.8B-details.MaziyarPanahi__calme-2.1-phi3.5-4b-details
Dataset Card for Evaluation run of MaziyarPanahi/calme-2.1-phi3.5-4b
Dataset automatically created during the evaluation run of model MaziyarPanahi/calme-2.1-phi3.5-4b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/MaziyarPanahi__calme-2.1-phi3.5-4b-details.MaziyarPanahi__calme-2.2-phi3-4b-details
Dataset Card for Evaluation run of MaziyarPanahi/calme-2.2-phi3-4b
Dataset automatically created during the evaluation run of model MaziyarPanahi/calme-2.2-phi3-4b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/MaziyarPanahi__calme-2.2-phi3-4b-details.
