CoolFace
9 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01llamafactory /tiny-supervised-datasettexttext-generationn<1K4 likes44k downloads2y agoHugging Face02jchenyu /t5_large_supervised_proportional_1MThis data set is created by randomly sampling 1M documents from the large supervised proportional mixture from the T5 repository. The code to produce this sampled dataset can be found here. tabular1M<n<10M0 likes140 downloads4y agoHugging Face03anti-ai /ViNLI-SimCSE-supervisedtextsentence-similarity100K<n<1M1 likes45 downloads3y agoHugging Face04anti-ai /ViNLI-Healthcare-supervisedgatedtexttext-classification1M<n<10M0 likes35 downloads1y agoHugging Face05anti-ai /ViNLI-Zalo-supervisedtextsentence-similarity10K<n<100K1 likes29 downloads2y agoHugging Face06anti-ai /ViNLI-SimCSE-supervised_v2textsentence-similarity100K<n<1M0 likes19 downloads2y agoHugging Face07insightful-stays /airbnb-reviews-supervisedtabular1K<n<10K0 likes6 downloads9mo agoHugging Face08Zhaoming213 /SupervisedFine-Tuning-unrestricted Introfuction This is a Supervised Fine-Tuning dataset.Filtering out common rejection logic, legal statements, moralizing, and other uncomfortable elements found in generative AI. If you need a pre-trained dataset, please go to:https://huggingface.co/datasets/Zhaoming213/Pretrain-unrestricted Filter keywords keywords_list = [ "我无法回答", "我无法给出", "我无法提供", "我不能提供", "我拒绝提供", "我不具备", "我不拥有", "作为一个AI", "作为一个 AI ", "作为AI", "作为语言", "作为大语言", "作为程序", "作为一款", "我没有个人"… See the full description on the dataset page: https://huggingface.co/datasets/Zhaoming213/SupervisedFine-Tuning-unrestricted.text100K<n<1M0 likes5 downloads6mo agoHugging Face09insightful-stays /airbnb-reviews-supervised-improvementstext1K<n<10K0 likes3 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.