datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
autotrain-data-chatxCNXT
/
CHaTx
likes
x #
Adapter Transformers
OpenAssistant/oasst1
fka/awesome-chatgpt-prompts
togethercomputer/RedPajama-Data-1T
anon8231489123/ShareGPT_Vicuna_unfiltered
gsdf/EasyNegative
bloomberg/entsum
openai/summarize_from_feedback
billsum
AmazonScience/massive
amazon_us_reviews
amazon_reviews_multi
openwebtext
microsoft/CLUES
Norod78/microsoft-fluentui-emoji-512-whitebg
Norod78/microsoft-fluentui-emoji-768
MicPie/unpredictable_msdn-microsoft-com
microsoft/codexglue_method_generation… See the full description on the dataset page: https://huggingface.co/datasets/CNXT/autotrain-data-chatx.mida-autotrain
SPIRIT Dataset (System Prompt Instruction Real-world Implementation Training-set)
Dataset Summary
SPIRIT is a high-quality system prompt instruction dataset designed to enhance language models' ability to follow complex system prompts. The dataset comprises real-world system prompts collected from GitHub repositories and synthetically generated conversations, specifically curated to improve system prompt adherence in large language models.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/t4t455/mida-autotrain.
