datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llama3.1_8b_inst_as_ver_gemma27b_it_math158_32gen_asyncfineweb-gemma27b-residualsdetails_google__gemma-3-27b-pt_v2
Dataset Card for Evaluation run of google/gemma-3-27b-pt
Dataset automatically created during the evaluation run of model google/gemma-3-27b-pt.
The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_google__gemma-3-27b-pt_v2.sorrel-T-gemma-3-27b-pt-seed0-documentsdetails_google__gemma-3-27b-it_v2
Dataset Card for Evaluation run of google/gemma-3-27b-it
Dataset automatically created during the evaluation run of model google/gemma-3-27b-it.
The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_google__gemma-3-27b-it_v2.gemma-3-27b-it-eval-logs-and-scorescrcis-quranic-eval-leaderboard-results_details_google__gemma-2-27b-it_private
Dataset Card for Evaluation run of google/gemma-2-27b-it
Dataset automatically created during the evaluation run of model google/gemma-2-27b-it.
The dataset is composed of 7 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/sadra-barikbin/crcis-quranic-eval-leaderboard-results_details_google__gemma-2-27b-it_private.wmt19-gemma-3-27b-it-CoTfineweb-gemma27b-texts-splittopxgen-gemma-3-27b-and-nllb-3.3b
TopXGen: Topic-Diverse Parallel Data for Low-Resource MT
Dataset Summary
This dataset is a synthetic parallel dataset for 10 low-resource languages, created by applying the TopXGen pipeline with recent multilingual LLMs. It is designed for machine translation (MT) fine-tuning and few-shot experiments (as a selection pool).The pipeline works as follows:
Topic-diverse paragraph generation in the target low-resource language using an LLM (generator), with diversity… See the full description on the dataset page: https://huggingface.co/datasets/almanach/topxgen-gemma-3-27b-and-nllb-3.3b.fineweb-gemma27b-embedswmt19-gemma-3-27b-it-SBYSmed-synth-questions-gemma-3-27b-it
openmed-community/med-synth-questions-gemma-3-27b-it
What is this?
Med Synth Questions — Gemma 3 27B IT is a compact, instruction-only dataset of 33,325 English medical questions generated from the texts in gamino/wiki_medical_terms. Each row contains a single question (input), the structured generation_settings used to produce it, and an ISO-8601 timestamp. The source corpus consists of Wikipedia-based medical term pages assembled in wiki_medical_terms (licensed GPL-3.0)… See the full description on the dataset page: https://huggingface.co/datasets/openmed-community/med-synth-questions-gemma-3-27b-it.wmt19-gemma-3-27b-it-DecompNAPS-ai__naps-gemma-2-27b-v-0.1.0-details
Dataset Card for Evaluation run of NAPS-ai/naps-gemma-2-27b-v-0.1.0
Dataset automatically created during the evaluation run of model NAPS-ai/naps-gemma-2-27b-v-0.1.0
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NAPS-ai__naps-gemma-2-27b-v-0.1.0-details.google__gemma-2-27b-it-details
Dataset Card for Evaluation run of google/gemma-2-27b-it
Dataset automatically created during the evaluation run of model google/gemma-2-27b-it
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/google__gemma-2-27b-it-details.med-synth-questions-gemma-3-27b-deepseek-v4-flash
Med Synth Questions (Gemma-3 + DeepSeek V4 Flash)
Synthetic reasoning traces and answers for medical questions from openmed-community/med-synth-questions-gemma-3-27b-it. Each record contains a medical question with SYNTH-style reasoning and a generated answer by DeepSeek V4 Flash.
Dataset Summary
29,148 records (2 dupes + 3,410 incomplete/truncated removed from 32,560 source)
29,148 reasoning turns (99.2% format compliance)
Average 1,591 chars per reasoning trace… See the full description on the dataset page: https://huggingface.co/datasets/mkurman/med-synth-questions-gemma-3-27b-deepseek-v4-flash.NAPS-ai__naps-gemma-2-27b-v0.1.0-details
Dataset Card for Evaluation run of NAPS-ai/naps-gemma-2-27b-v0.1.0
Dataset automatically created during the evaluation run of model NAPS-ai/naps-gemma-2-27b-v0.1.0
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NAPS-ai__naps-gemma-2-27b-v0.1.0-details.fineweb-gemma27b-textswmt19-gemma-3-27b-it-TEaRtopxgen-gemma-3-27b-it-SBYSINSAIT-Institute__BgGPT-Gemma-2-27B-IT-v1.0-details
Dataset Card for Evaluation run of INSAIT-Institute/BgGPT-Gemma-2-27B-IT-v1.0
Dataset automatically created during the evaluation run of model INSAIT-Institute/BgGPT-Gemma-2-27B-IT-v1.0
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/INSAIT-Institute__BgGPT-Gemma-2-27B-IT-v1.0-details.wmt19-gemma-3-27b-it-MAPSdetails_google__gemma-3-27b-it_v2_alrage
Dataset Card for Evaluation run of google/gemma-3-27b-it
Dataset automatically created during the evaluation run of model google/gemma-3-27b-it.
The dataset is composed of 1 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_google__gemma-3-27b-it_v2_alrage.topxgen-gemma-3-27b-and-nllb-3.3b
TopXGen: Topic-Diverse Parallel Data for Low-Resource MT
Dataset Summary
This dataset is a synthetic parallel dataset for 10 low-resource languages, created by applying the TopXGen pipeline with recent multilingual LLMs. It is designed for machine translation (MT) fine-tuning and few-shot experiments (as a selection pool).The pipeline works as follows:
Topic-diverse paragraph generation in the target low-resource language using an LLM (generator), with diversity… See the full description on the dataset page: https://huggingface.co/datasets/innadark/topxgen-gemma-3-27b-and-nllb-3.3b.GCM-dataset-gemma-3-27b-it
wikipedia_clean-gemma-3-27b-it-T0.0details_byroneverson__gemma-2-27b-it-abliterated
Dataset Card for Evaluation run of byroneverson/gemma-2-27b-it-abliterated
Dataset automatically created during the evaluation run of model byroneverson/gemma-2-27b-it-abliterated.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_byroneverson__gemma-2-27b-it-abliterated.GCM-dataset-TriviaQA_gemma-3-27b-it
details_migtissera__Tess-v2.5-Gemma-2-27B-alpha
Dataset Card for Evaluation run of migtissera/Tess-v2.5-Gemma-2-27B-alpha
Dataset automatically created during the evaluation run of model migtissera/Tess-v2.5-Gemma-2-27B-alpha.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_migtissera__Tess-v2.5-Gemma-2-27B-alpha.
