CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01aumghag /Data-Analytics-Digital-Marketing-Project-Management-QA_DBtextquestion-answeringn<1K4 likes135 downloads2y agoHugging Face02bitext /Bitext-wealth-management-llm-chatbot-training-dataset Bitext - Wealth Management Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the [Wealth Management] sector can be easily achieved using our two-step approach to LLM… See the full description on the dataset page: https://huggingface.co/datasets/bitext/Bitext-wealth-management-llm-chatbot-training-dataset.textquestion-answering10K<n<100K2 likes90 downloads2y agoHugging Face03avery-ma /ManyHarmgated Dataset Card for ManyHarm Paper: PANDAS: Improving Many-shot Jailbreaking via Positive Affirmation, Negative Demonstration, and Adaptive Sampling 🔄 Update August 2, 2025: We have observed a growing number of access requests from accounts using temporary or disposable email providers. To ensure responsible use and maintain the integrity of our access policy, requests from such accounts will be denied. We recommend using a valid, verifiable institutional or… See the full description on the dataset page: https://huggingface.co/datasets/avery-ma/ManyHarm.textquestion-answering1K<n<10K0 likes25 downloads1y agoHugging Face04ManjuKrish /llm-delusion-response-annotations LLM Delusion-Like Belief Reinforcement Annotations This dataset contains human annotations of responses generated by conversational large language models (LLMs) to prompts expressing potentially delusion-like or reality-distorted beliefs. The purpose of the dataset is to support evaluation of whether conversational LLM responses may unintentionally reinforce or strengthen delusion-like beliefs. Dataset Files Consensus Dataset… See the full description on the dataset page: https://huggingface.co/datasets/ManjuKrish/llm-delusion-response-annotations.documenttext-classification1K<n<10K0 likes25 downloads3mo agoHugging Face05manojroyal23 /customer-support-tickets Featuring Labeled Customer Emails and Support Responses 🔧 Synthetic IT Ticket Generator — Custom Dataset Create a dataset tailored to your own queues & priorities (no PII). 👉 Generate custom data Define your queues, priorities, language Need an on-prem AI to auto-classify tickets?→ Open Ticket AI There are 2 Versions of the dataset, the new version has more tickets, but only languages english and german. So please look at both files, to find what best fits your needs.… See the full description on the dataset page: https://huggingface.co/datasets/manojroyal23/customer-support-tickets.texttext-classification10K<n<100K0 likes23 downloads6mo agoHugging Face06datasetter458 /bash-reference-manual-general-QAs Dataset generated from bash reference manual. book information like date and bash version are available within the very first rows of the dataset this dataset is pretty small in general, but covering almost all of the definition and technical terms, commands and flags in the book columns : "Question", "Answer" texttext-generation1K<n<10K0 likes15 downloads4mo agoHugging Face075digit /Many-of-you-have-heard-that-intexttext-classification1K<n<10K0 likes12 downloads1y agoHugging Face08manojbaniya /ecommerce_qna Ecommerce Customer Query Dataset Roman Nepali Dataset Generated for Customer Support Question Answer textquestion-answering10K<n<100K2 likes7 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.