CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MugenYume /OpenHermes-2.5-Autotrain-SFT This is the converted OpenHermes 2.5 dataset, available here: teknium/OpenHermes-2.5 All credit goes to teknium for creating the original dataset. This version has been specifically formatted for training large language models (LLMs) using HuggingFace AutoTrain. The dataset now contains a single text column, optimized for the LLM SFT training method. You can find other versions of the dataset in my repository as well. I have filtered the dataset in various ways. For example, if you're not… See the full description on the dataset page: https://huggingface.co/datasets/MugenYume/OpenHermes-2.5-Autotrain-SFT.text1M<n<10M0 likes34 downloads2y agoHugging Face02open-llm-leaderboard /trthminh1112__autotrain-llama32-1b-finetune-detailsgated Dataset Card for Evaluation run of trthminh1112/autotrain-llama32-1b-finetune Dataset automatically created during the evaluation run of model trthminh1112/autotrain-llama32-1b-finetune The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/trthminh1112__autotrain-llama32-1b-finetune-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face03wesley7137 /autotrain_qa_neurotext1K<n<10K0 likes23 downloads3y agoHugging Face04ahmed000000000 /autotrain-data-test1text10K<n<100K0 likes22 downloads3y agoHugging Face05lucasbrandao /autotrain-data-llama-autotraintextn<1K0 likes15 downloads3y agoHugging Face06Kasd007 /auto-train-1-1776291068490text1K<n<10K0 likes14 downloads5mo agoHugging Face07open-llm-leaderboard /abhishek__autotrain-llama3-70b-orpo-v2-detailsgated Dataset Card for Evaluation run of abhishek/autotrain-llama3-70b-orpo-v2 Dataset automatically created during the evaluation run of model abhishek/autotrain-llama3-70b-orpo-v2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abhishek__autotrain-llama3-70b-orpo-v2-details.tabular10K<n<100K0 likes13 downloads2y agoHugging Face08asrini27 /autotrain-data-test-v2textn<1K0 likes12 downloads1y agoHugging Face09meirm /autotrain-poumpourastext1K<n<10K0 likes11 downloads2y agoHugging Face10shashwat16 /autotrain-data-dialect-translationn<1K0 likes10 downloads3y agoHugging Face11ogiwemy /autotrain-data-rekomendasi-mata-kuliahtextn<1K0 likes9 downloads3y agoHugging Face12Kasd007 /auto-train-1-1776353045609text1K<n<10K0 likes9 downloads5mo agoHugging Face13Akhil333 /autotrain-data-ruzrttext100K<n<1M0 likes8 downloads3y agoHugging Face14open-llm-leaderboard /abhishek__autotrain-llama3-orpo-v2-detailsgated Dataset Card for Evaluation run of abhishek/autotrain-llama3-orpo-v2 Dataset automatically created during the evaluation run of model abhishek/autotrain-llama3-orpo-v2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abhishek__autotrain-llama3-orpo-v2-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face15open-llm-leaderboard /abhishek__autotrain-llama3-70b-orpo-v1-detailsgated Dataset Card for Evaluation run of abhishek/autotrain-llama3-70b-orpo-v1 Dataset automatically created during the evaluation run of model abhishek/autotrain-llama3-70b-orpo-v1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abhishek__autotrain-llama3-70b-orpo-v1-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face16open-llm-leaderboard /abhishek__autotrain-vr4a1-e5mms-detailsgated Dataset Card for Evaluation run of abhishek/autotrain-vr4a1-e5mms Dataset automatically created during the evaluation run of model abhishek/autotrain-vr4a1-e5mms The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abhishek__autotrain-vr4a1-e5mms-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face17RandomOscillations /autotrain-data-amdal-mining-llama2-7b-cleantabularn<1K0 likes8 downloads1y agoHugging Face18semanticword-user /autotrain-dataset-2textn<1K0 likes7 downloads3y agoHugging Face19Battmanux /autotrain-1o6jf-k6pnmtextn<1K0 likes7 downloads2y agoHugging Face20open-llm-leaderboard /abhishek__autotrain-0tmgq-5tpbg-detailsgated Dataset Card for Evaluation run of abhishek/autotrain-0tmgq-5tpbg Dataset automatically created during the evaluation run of model abhishek/autotrain-0tmgq-5tpbg The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abhishek__autotrain-0tmgq-5tpbg-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face21Kasd007 /auto-train-1-1774182339942text10K<n<100K0 likes7 downloads6mo agoHugging Face22Kasd007 /auto-train-1-1774220619885text10K<n<100K0 likes7 downloads6mo agoHugging Face23t4t455 /mida-autotrain2 RuTurboAlpaca Dataset of ChatGPT-generated instructions in Russian. Code: rulm/self_instruct Code is based on Stanford Alpaca and self-instruct. 29822 examples Preliminary evaluation by an expert based on 400 samples: 83% of samples contain correct instructions 63% of samples have correct instructions and outputs Crowdsouring-based evaluation on 3500 samples: 90% of samples contain correct instructions 68% of samples have correct instructions and outputs Prompt template:… See the full description on the dataset page: https://huggingface.co/datasets/t4t455/mida-autotrain2.tabulartext-generation10K<n<100K0 likes7 downloads3mo agoHugging Face24Mockup /autotrain-data-writingtext10K<n<100K1 likes6 downloads4y agoHugging Face25HsiangNianian /autotrain-data-chinese-nertexttext-generationn<1K0 likes6 downloads3y agoHugging Face26heinsithu /autotrain-data-52y0-943s-i8cytextn<1K0 likes6 downloads3y agoHugging Face27Kasd007 /auto-train-1-1774285994483text1K<n<10K0 likes6 downloads6mo agoHugging Face28t4t455 /mida-autotrain SPIRIT Dataset (System Prompt Instruction Real-world Implementation Training-set) Dataset Summary SPIRIT is a high-quality system prompt instruction dataset designed to enhance language models' ability to follow complex system prompts. The dataset comprises real-world system prompts collected from GitHub repositories and synthetically generated conversations, specifically curated to improve system prompt adherence in large language models. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/t4t455/mida-autotrain.textquestion-answering10K<n<100K0 likes5 downloads3mo agoHugging Face29thatdudejbob /autotrain-data-uo-codetext1K<n<10K0 likes4 downloads2y agoHugging Face30mgamal /autotrain-data-address-parsingtabularn<1K0 likes4 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.