datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Mistral-7B-v0.1-base-tokenized-dolma-v1_7-50Bmistral-7b-hidden-states-tpu-verification
Mistral-7B TPU hidden-state extractor verification
Verification-only artifact — not a reproduction of the paper's AUROC.
This dataset contains last-token embeddings and hidden states extracted from
1,000 flattened CoQA validation question/reference-answer pairs with
mistralai/Mistral-7B-Instruct-v0.3. It verifies that the memory-bounded TPU
v5e-1 extraction path runs successfully. It does not contain the Mistral
best_answer generations or hallucination labels required to… See the full description on the dataset page: https://huggingface.co/datasets/GwendalTsang/mistral-7b-hidden-states-tpu-verification.Mistral-7B-LongPO-256K-tokenizedMistral-7B-LongPO-128K-tokenizeddetails_xxyyy123__Mistral7B_adaptor_v1
Dataset Card for Evaluation run of xxyyy123/Mistral7B_adaptor_v1
Dataset Summary
Dataset automatically created during the evaluation run of model xxyyy123/Mistral7B_adaptor_v1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_xxyyy123__Mistral7B_adaptor_v1.details_RefalMachine__ruadapt_mistral7b_full_vo_1e4Mistral-7B-LongPO-512K-tokenizedlm-eval-results-chlee10-T3Q-Merge-Mistral7B-private
Dataset Card for Evaluation run of chlee10/T3Q-Merge-Mistral7B
Dataset automatically created during the evaluation run of model chlee10/T3Q-Merge-Mistral7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-chlee10-T3Q-Merge-Mistral7B-private.details_chlee10__T3Q-Merge-Mistral7B
Dataset Card for Evaluation run of chlee10/T3Q-Merge-Mistral7B
Dataset automatically created during the evaluation run of model chlee10/T3Q-Merge-Mistral7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_chlee10__T3Q-Merge-Mistral7B.Mistral-7B-v0.1-base-tokenized-fineweb-edu-45B-4096openr1-mistral7b-fulldetails_chlee10__T3Q-Platypus-Mistral7B
Dataset Card for Evaluation run of chlee10/T3Q-Platypus-Mistral7B
Dataset automatically created during the evaluation run of model chlee10/T3Q-Platypus-Mistral7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_chlee10__T3Q-Platypus-Mistral7B.mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_160k_ratiomp_mistral7bv3_sft_ogd_rms_epoch5_40k_multisample_n2mpg27_mistral7bv3_sft_multisample_2.5kmp_mistral7bv3_sft_dpo_beta5e-2_epoch1_multisample_2.5knomiracl_mistral7B_filter_langs
Dataset Card for "nomiracl_mistral7B_filter_langs"
More Information needed
mp_mistral7bv3_sft_dpo_beta1e-1_epoch1_40k_n16mpg27_mistral7bv3_sft_dpo_beta5e-2_epoch1_40k_multisample_ratiomp_mistral7bv3_sft_dpo_beta2e-2_epoch1_20k_n8Mistral-7b-0.3-Instruct-TriviaQA-HighlyKnownDataset for paper “How Much Knowledge Can You Pack into a LoRA Adapter without Harming LLM?”
Based on TriviaQA dataset
(https://huggingface.co/papers/2502.14502)
mp_mistral7bv3_sft_40k_multisample_n4mistral-7b-arxiv-paper-chunkedThis dataset contains chunked extracts from the Mistral 7B research paper.
mistral-7b-utf-self-judgeMistral-7b-0.3-Instruct-DBpedia-HighlyKnownDataset for paper “How Much Knowledge Can You Pack into a LoRA Adapter without Harming LLM?”
Based on DBpedia dataset
Paper, Code
mp_mistral7bv3_sft_dpo_beta5e-2_epoch1_40kUCLA-AGI__Mistral7B-PairRM-SPPOmp_mistral7bv3_sft_ogd_rms_epoch5_40kmpg27_mistral7bv3_sft_dpo_beta2e-1_epoch2_40kMedical-QA-Mistral7B-Finetuning
