datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
LDS-retrain-bank-adamw-N16k-bs256-gpt2-mediumPARTIAL_LDS-retrain-bank-gpt2medium-16k-bs32details_LordNoah__latent_gpt2_medium_alpaca_e2
Dataset Card for Evaluation run of LordNoah/latent_gpt2_medium_alpaca_e2
Dataset automatically created during the evaluation run of model LordNoah/latent_gpt2_medium_alpaca_e2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LordNoah__latent_gpt2_medium_alpaca_e2.details_postbot__gpt2-medium-emailgen
Dataset Card for Evaluation run of postbot/gpt2-medium-emailgen
Dataset Summary
Dataset automatically created during the evaluation run of model postbot/gpt2-medium-emailgen on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_postbot__gpt2-medium-emailgen.gpt2_fineweb_kl_medium_largegpt2_openwebtext_kl_medium_xldetails_LordNoah__latent_gpt2_medium_alpaca_e3
Dataset Card for Evaluation run of LordNoah/latent_gpt2_medium_alpaca_e3
Dataset automatically created during the evaluation run of model LordNoah/latent_gpt2_medium_alpaca_e3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LordNoah__latent_gpt2_medium_alpaca_e3.details_LordNoah__spin_gpt2_medium_alpaca_e3
Dataset Card for Evaluation run of LordNoah/spin_gpt2_medium_alpaca_e3
Dataset automatically created during the evaluation run of model LordNoah/spin_gpt2_medium_alpaca_e3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LordNoah__spin_gpt2_medium_alpaca_e3.details_LordNoah__latent_gpt2_medium_alpaca_e4
Dataset Card for Evaluation run of LordNoah/latent_gpt2_medium_alpaca_e4
Dataset automatically created during the evaluation run of model LordNoah/latent_gpt2_medium_alpaca_e4 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LordNoah__latent_gpt2_medium_alpaca_e4.details_LordNoah__spin_gpt2_medium_alpaca_e2
Dataset Card for Evaluation run of LordNoah/spin_gpt2_medium_alpaca_e2
Dataset automatically created during the evaluation run of model LordNoah/spin_gpt2_medium_alpaca_e2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_LordNoah__spin_gpt2_medium_alpaca_e2.gpt2_fineweb_kl_small_mediumgpt2_slimpajama_kl_small_mediumsquad_30_percent_pruned_by_ppl_gpt2-medium
Dataset Card for "squad_30_percent_pruned_by_ppl_gpt2-medium"
More Information needed
shaer-eval-raw-gpt2-medium-arabic-poetry
Raw Shaer Continuation Generations - gpt2_medium_arabic_poetry
This dataset contains cumulative raw continuation generations for gpt2_medium_arabic_poetry.
The main table is data/test.jsonl; it includes source/prompt/reference fields plus the model output in generated_text and raw_generated_text.
Current uploaded progress target: 3481 rows out of 3481.
Scored upload: 1. When scored upload is true, reward/evaluation columns are included in the main table.
Artifacts such as… See the full description on the dataset page: https://huggingface.co/datasets/Shaer-AI-2/shaer-eval-raw-gpt2-medium-arabic-poetry.gpt2-medium_dpo_anthropic_hhppl_gpt2-medium_ranked_squad
Dataset Card for "ppl_gpt2-medium_ranked_squad"
More Information needed
openai-community__gpt2-medium-details
Dataset Card for Evaluation run of openai-community/gpt2-medium
Dataset automatically created during the evaluation run of model openai-community/gpt2-medium
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/openai-community__gpt2-medium-details.gpt2_c4_kl_medium_largegpt2_c4_kl_small_mediummkd-1-4-gpt2-mediumtldr_gpt2_medium_w2s_feedbackhh_gpt2_medium_w2s_feedbackgpt2_c4_kl_medium_xlgpt2-medium-solutionsgpt2_pile_kl_medium_xlgpt2-medium_dpo_tldr_temp_1_2
Dataset Card for "gpt2-medium_dpo_tldr_temp_1_2"
More Information needed
wikitext-2-v1-kl-gpt2-medium-vs-gpt2-largeEve-casual-gpt2-mediumweak_gpt2-medium_tldr_syntheticweak_gpt2_medium_dpo_hh
Dataset Card for "weak_gpt2_medium_dpo_hh"
More Information needed
