datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
monarchwikipedia
Dataset Card for Wikimedia Wikipedia
Dataset Summary
Wikipedia dataset containing cleaned articles of all languages.
The dataset is built from the Wikipedia dumps (https://dumps.wikimedia.org/)
with one subset per language, each containing a single train split.
Each example contains the content of one full Wikipedia article with cleaning to strip
markdown and unwanted sections (references, etc.).
All language subsets have already been processed for recent dump… See the full description on the dataset page: https://huggingface.co/datasets/Monarch700/wikipedia.details_macadeliccc__Monarch-7B-SFT
Dataset Card for Evaluation run of macadeliccc/Monarch-7B-SFT
Dataset automatically created during the evaluation run of model macadeliccc/Monarch-7B-SFT on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_macadeliccc__Monarch-7B-SFT.lm-eval-results-AtAndDev-Ogno-Monarch-Neurotic-7B-Dare-Ties-private
Dataset Card for Evaluation run of AtAndDev/Ogno-Monarch-Neurotic-7B-Dare-Ties
Dataset automatically created during the evaluation run of model AtAndDev/Ogno-Monarch-Neurotic-7B-Dare-Ties
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-AtAndDev-Ogno-Monarch-Neurotic-7B-Dare-Ties-private.details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test.details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO.dragon-ai-vector-embeddingsdetails_ichigoberry__MonarchPipe-7B-slerp
Dataset Card for Evaluation run of ichigoberry/MonarchPipe-7B-slerp
Dataset automatically created during the evaluation run of model ichigoberry/MonarchPipe-7B-slerp on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ichigoberry__MonarchPipe-7B-slerp.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-private.details_giraffe176__Open_Neural_Monarch_Maidv0.1
Dataset Card for Evaluation run of giraffe176/Open_Neural_Monarch_Maidv0.1
Dataset automatically created during the evaluation run of model giraffe176/Open_Neural_Monarch_Maidv0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Neural_Monarch_Maidv0.1.C3PO
CHEBI Chemical Classification Program Ontology (C3PO) Benchmark
Name: C3PO-v237
A benchmark for chemical classification, derived from v237 of the CHEBI ontology.
The benchmark consists of
177,875 Structures, each represented by a SMILES string
1364 Classes, each of which classify (directly or indirectly) between 25 and 5000 Structures
The CHEBI smiles property (obo:chebi/smiles) is used to derive
structures. Only SMILES strings that lacked a * wildcard are
used. If multiple CHEBI… See the full description on the dataset page: https://huggingface.co/datasets/MonarchInit/C3PO.details_eren23__ogno-monarch-jaskier-merge-7b
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_eren23__ogno-monarch-jaskier-merge-7b.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-v2-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-v2
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-v2
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-v2-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test-private.imnet1k_monarch_monarch_butterfly_milkweed_butterfly_Danaus_plexippuslm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2-private.details_abideen__MonarchCoder-MoE-2x7B
Dataset Card for Evaluation run of abideen/MonarchCoder-MoE-2x7B
Dataset automatically created during the evaluation run of model abideen/MonarchCoder-MoE-2x7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abideen__MonarchCoder-MoE-2x7B.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3-private.details_mlabonne__Monarch-7B
Dataset Card for Evaluation run of mlabonne/Monarch-7B
Dataset automatically created during the evaluation run of model mlabonne/Monarch-7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_mlabonne__Monarch-7B.details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2.details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_eren23__ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3.details_abideen__MonarchCoder-7B
Dataset Card for Evaluation run of abideen/MonarchCoder-7B
Dataset automatically created during the evaluation run of model abideen/MonarchCoder-7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abideen__MonarchCoder-7B.details_AtAndDev__Ogno-Monarch-Neurotic-7B-Dare-Ties
Dataset Card for Evaluation run of AtAndDev/Ogno-Monarch-Neurotic-7B-Dare-Ties
Dataset automatically created during the evaluation run of model AtAndDev/Ogno-Monarch-Neurotic-7B-Dare-Ties on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_AtAndDev__Ogno-Monarch-Neurotic-7B-Dare-Ties.details_giraffe176__Starling_Monarch_Westlake_Garten-7B-v0.1
Dataset Card for Evaluation run of giraffe176/Starling_Monarch_Westlake_Garten-7B-v0.1
Dataset automatically created during the evaluation run of model giraffe176/Starling_Monarch_Westlake_Garten-7B-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Starling_Monarch_Westlake_Garten-7B-v0.1.details_eren23__ogno-monarch-jaskier-merge-7b-v2
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-v2
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-v2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_eren23__ogno-monarch-jaskier-merge-7b-v2.wan_fewstep_monarch_fast_framewisedetails_macadeliccc__MonarchLake-7B
Dataset Card for Evaluation run of macadeliccc/MonarchLake-7B
Dataset automatically created during the evaluation run of model macadeliccc/MonarchLake-7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_macadeliccc__MonarchLake-7B.details_AtAndDev__Ogno-Monarch-Neurotic-9B-Passthrough
Dataset Card for Evaluation run of AtAndDev/Ogno-Monarch-Neurotic-9B-Passthrough
Dataset automatically created during the evaluation run of model AtAndDev/Ogno-Monarch-Neurotic-9B-Passthrough on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_AtAndDev__Ogno-Monarch-Neurotic-9B-Passthrough.details_giraffe176__Open_Neural_Monarch_Maidv0.2
Dataset Card for Evaluation run of giraffe176/Open_Neural_Monarch_Maidv0.2
Dataset automatically created during the evaluation run of model giraffe176/Open_Neural_Monarch_Maidv0.2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Neural_Monarch_Maidv0.2.
