CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ReasoningMila /llama3.1_8b_inst_as_ver_gemma27b_it_math158_32gen_async0 likes2k downloads2y agoHugging Face02nickypro /fineweb-gemma27b-residuals0 likes339 downloads11mo agoHugging Face03OALL /details_google__gemma-3-27b-pt_v2 Dataset Card for Evaluation run of google/gemma-3-27b-pt Dataset automatically created during the evaluation run of model google/gemma-3-27b-pt. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_google__gemma-3-27b-pt_v2.text100K<n<1M0 likes194 downloads1y agoHugging Face04joshycodes /sorrel-T-gemma-3-27b-pt-seed0-documentstext100K<n<1M0 likes179 downloads8d agoHugging Face05OALL /details_google__gemma-3-27b-it_v2 Dataset Card for Evaluation run of google/gemma-3-27b-it Dataset automatically created during the evaluation run of model google/gemma-3-27b-it. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_google__gemma-3-27b-it_v2.text100K<n<1M0 likes177 downloads1y agoHugging Face06twinkle-ai /gemma-3-27b-it-eval-logs-and-scorestabular100K<n<1M0 likes164 downloads7mo agoHugging Face07sadra-barikbin /crcis-quranic-eval-leaderboard-results_details_google__gemma-2-27b-it_private Dataset Card for Evaluation run of google/gemma-2-27b-it Dataset automatically created during the evaluation run of model google/gemma-2-27b-it. The dataset is composed of 7 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/sadra-barikbin/crcis-quranic-eval-leaderboard-results_details_google__gemma-2-27b-it_private.textn<1K0 likes122 downloads2y agoHugging Face08almanach /wmt19-gemma-3-27b-it-CoTtext100K<n<1M0 likes119 downloads1y agoHugging Face09annnettte /fineweb-gemma27b-texts-splittext1M<n<10M0 likes114 downloads1y agoHugging Face10almanach /topxgen-gemma-3-27b-and-nllb-3.3b TopXGen: Topic-Diverse Parallel Data for Low-Resource MT Dataset Summary This dataset is a synthetic parallel dataset for 10 low-resource languages, created by applying the TopXGen pipeline with recent multilingual LLMs. It is designed for machine translation (MT) fine-tuning and few-shot experiments (as a selection pool).The pipeline works as follows: Topic-diverse paragraph generation in the target low-resource language using an LLM (generator), with diversity… See the full description on the dataset page: https://huggingface.co/datasets/almanach/topxgen-gemma-3-27b-and-nllb-3.3b.texttranslation1M<n<10M4 likes98 downloads1y agoHugging Face11nickypro /fineweb-gemma27b-embeds0 likes74 downloads10mo agoHugging Face12almanach /wmt19-gemma-3-27b-it-SBYStext100K<n<1M0 likes72 downloads1y agoHugging Face13openmed-community /med-synth-questions-gemma-3-27b-it openmed-community/med-synth-questions-gemma-3-27b-it What is this? Med Synth Questions — Gemma 3 27B IT is a compact, instruction-only dataset of 33,325 English medical questions generated from the texts in gamino/wiki_medical_terms. Each row contains a single question (input), the structured generation_settings used to produce it, and an ISO-8601 timestamp. The source corpus consists of Wikipedia-based medical term pages assembled in wiki_medical_terms (licensed GPL-3.0)… See the full description on the dataset page: https://huggingface.co/datasets/openmed-community/med-synth-questions-gemma-3-27b-it.texttext-generation10K<n<100K4 likes63 downloads1y agoHugging Face14almanach /wmt19-gemma-3-27b-it-Decomptext100K<n<1M0 likes58 downloads1y agoHugging Face15open-llm-leaderboard /NAPS-ai__naps-gemma-2-27b-v-0.1.0-detailsgated Dataset Card for Evaluation run of NAPS-ai/naps-gemma-2-27b-v-0.1.0 Dataset automatically created during the evaluation run of model NAPS-ai/naps-gemma-2-27b-v-0.1.0 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NAPS-ai__naps-gemma-2-27b-v-0.1.0-details.tabular10K<n<100K0 likes50 downloads2y agoHugging Face16open-llm-leaderboard /google__gemma-2-27b-it-detailsgated Dataset Card for Evaluation run of google/gemma-2-27b-it Dataset automatically created during the evaluation run of model google/gemma-2-27b-it The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/google__gemma-2-27b-it-details.tabular10K<n<100K0 likes49 downloads2y agoHugging Face17mkurman /med-synth-questions-gemma-3-27b-deepseek-v4-flash Med Synth Questions (Gemma-3 + DeepSeek V4 Flash) Synthetic reasoning traces and answers for medical questions from openmed-community/med-synth-questions-gemma-3-27b-it. Each record contains a medical question with SYNTH-style reasoning and a generated answer by DeepSeek V4 Flash. Dataset Summary 29,148 records (2 dupes + 3,410 incomplete/truncated removed from 32,560 source) 29,148 reasoning turns (99.2% format compliance) Average 1,591 chars per reasoning trace… See the full description on the dataset page: https://huggingface.co/datasets/mkurman/med-synth-questions-gemma-3-27b-deepseek-v4-flash.tabulartext-generation10K<n<100K1 likes49 downloads2mo agoHugging Face18open-llm-leaderboard /NAPS-ai__naps-gemma-2-27b-v0.1.0-detailsgated Dataset Card for Evaluation run of NAPS-ai/naps-gemma-2-27b-v0.1.0 Dataset automatically created during the evaluation run of model NAPS-ai/naps-gemma-2-27b-v0.1.0 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NAPS-ai__naps-gemma-2-27b-v0.1.0-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face19annnettte /fineweb-gemma27b-textstext1M<n<10M0 likes44 downloads1y agoHugging Face20almanach /wmt19-gemma-3-27b-it-TEaRtext100K<n<1M0 likes44 downloads1y agoHugging Face21almanach /topxgen-gemma-3-27b-it-SBYStext100K<n<1M0 likes43 downloads1y agoHugging Face22open-llm-leaderboard /INSAIT-Institute__BgGPT-Gemma-2-27B-IT-v1.0-detailsgated Dataset Card for Evaluation run of INSAIT-Institute/BgGPT-Gemma-2-27B-IT-v1.0 Dataset automatically created during the evaluation run of model INSAIT-Institute/BgGPT-Gemma-2-27B-IT-v1.0 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/INSAIT-Institute__BgGPT-Gemma-2-27B-IT-v1.0-details.tabular10K<n<100K0 likes41 downloads2y agoHugging Face23almanach /wmt19-gemma-3-27b-it-MAPStext100K<n<1M0 likes40 downloads1y agoHugging Face24OALL /details_google__gemma-3-27b-it_v2_alrage Dataset Card for Evaluation run of google/gemma-3-27b-it Dataset automatically created during the evaluation run of model google/gemma-3-27b-it. The dataset is composed of 1 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_google__gemma-3-27b-it_v2_alrage.text1K<n<10K0 likes39 downloads1y agoHugging Face25innadark /topxgen-gemma-3-27b-and-nllb-3.3b TopXGen: Topic-Diverse Parallel Data for Low-Resource MT Dataset Summary This dataset is a synthetic parallel dataset for 10 low-resource languages, created by applying the TopXGen pipeline with recent multilingual LLMs. It is designed for machine translation (MT) fine-tuning and few-shot experiments (as a selection pool).The pipeline works as follows: Topic-diverse paragraph generation in the target low-resource language using an LLM (generator), with diversity… See the full description on the dataset page: https://huggingface.co/datasets/innadark/topxgen-gemma-3-27b-and-nllb-3.3b.texttranslation1M<n<10M0 likes36 downloads9mo agoHugging Face26Hanqix /GCM-dataset-gemma-3-27b-it text10K<n<100K0 likes31 downloads11d agoHugging Face27ArmelR /wikipedia_clean-gemma-3-27b-it-T0.0text1K<n<10K0 likes31 downloads2mo agoHugging Face28OALL /details_byroneverson__gemma-2-27b-it-abliterated Dataset Card for Evaluation run of byroneverson/gemma-2-27b-it-abliterated Dataset automatically created during the evaluation run of model byroneverson/gemma-2-27b-it-abliterated. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_byroneverson__gemma-2-27b-it-abliterated.tabular100K<n<1M0 likes30 downloads2y agoHugging Face29Hanqix /GCM-dataset-TriviaQA_gemma-3-27b-it text10K<n<100K0 likes29 downloads11d agoHugging Face30OALL /details_migtissera__Tess-v2.5-Gemma-2-27B-alpha Dataset Card for Evaluation run of migtissera/Tess-v2.5-Gemma-2-27B-alpha Dataset automatically created during the evaluation run of model migtissera/Tess-v2.5-Gemma-2-27B-alpha. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_migtissera__Tess-v2.5-Gemma-2-27B-alpha.tabular100K<n<1M0 likes28 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.