datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wikipedia
Dataset Card for Wikimedia Wikipedia
Dataset Summary
Wikipedia dataset containing cleaned articles of all languages.
The dataset is built from the Wikipedia dumps (https://dumps.wikimedia.org/)
with one subset per language, each containing a single train split.
Each example contains the content of one full Wikipedia article with cleaning to strip
markdown and unwanted sections (references, etc.).
All language subsets have already been processed for recent dump… See the full description on the dataset page: https://huggingface.co/datasets/Monarch700/wikipedia.lm-eval-results-AtAndDev-Ogno-Monarch-Neurotic-7B-Dare-Ties-private
Dataset Card for Evaluation run of AtAndDev/Ogno-Monarch-Neurotic-7B-Dare-Ties
Dataset automatically created during the evaluation run of model AtAndDev/Ogno-Monarch-Neurotic-7B-Dare-Ties
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-AtAndDev-Ogno-Monarch-Neurotic-7B-Dare-Ties-private.dragon-ai-vector-embeddingslm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-v2-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-v2
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-v2
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-v2-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v4-test-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v2-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-private.lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3-private
Dataset Card for Evaluation run of eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3
Dataset automatically created during the evaluation run of model eren23/ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-eren23-ogno-monarch-jaskier-merge-7b-OH-PREF-DPO-v3-private.miriad-5.8M
Dataset Summary
MIRIAD is a curated million scale Medical Instruction and RetrIeval Dataset. It contains 5.8 million medical question-answer pairs, distilled from peer-reviewed biomedical literature using LLMs. MIRIAD provides structured, high-quality QA pairs, enabling diverse downstream tasks like RAG, medical retrieval, hallucination detection, and instruction tuning.
The dataset was introduced in our arXiv preprint.
To load the dataset, run:
from datasets… See the full description on the dataset page: https://huggingface.co/datasets/MONARCH4842/miriad-5.8M.dragon-ai-definition-evalsResults of expert evaluation on definitions generated by LLMs and ontology editors
See: https://github.com/monarch-initiative/dragon-ai-results
Although this dataset is partially prediction results, the expert evaluations form a dataset that could be used for new AI
tasks, specifically: Can we use AI to predict which definitions are accurate, concise, consistent, etc?
monarch_embeddingsSee here for a jupyter notebook used to produce these embeddings:
https://github.com/justaddcoffee/embed_monarch
dataset:
name: "Monarch KG"
url: "https://data.monarchinitiative.org/monarch-kg/2024-02-13/monarch-kg.tar.gz"
title: "Monarch Knowledge Graph"
source: "Monarch Initiative"
version: "2024-02-13"
embedding_model:
name: "First-order LINE"
title: "First-order LINE (Large-scale Information Network Embedding) from the GRAPE implementation"
source: "GRAPE"… See the full description on the dataset page: https://huggingface.co/datasets/justaddcoffee/monarch_embeddings.
