datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SPML_Chatbot_Prompt_Injection
SPML Chatbot Prompt Injection Dataset
Arxiv Paper
Introducing the SPML Chatbot Prompt Injection Dataset: a robust collection of system prompts designed to create realistic chatbot interactions, coupled with a diverse array of annotated user prompts that attempt to carry out prompt injection attacks. While other datasets in this domain have centered on less practical chatbot scenarios or have limited themselves to "jailbreaking" – just one aspect of prompt injection – our dataset… See the full description on the dataset page: https://huggingface.co/datasets/reshabhs/SPML_Chatbot_Prompt_Injection.chatbot-arena-elo
LMSYS Chatbot Arena ELO Scores
This dataset is a datasets-friendly version of Chatbot Arena ELO scores,
updated daily from the leaderboard API at
https://huggingface.co/spaces/lmarena-ai/chatbot-arena-leaderboard.
Updated: 20250717
Loading Data
from datasets import load_dataset
dataset = load_dataset("mathewhe/chatbot-arena-elo", split="train")
The main branch of this dataset will always be updated to the latest ELO and
leaderboard version. If you need a fixed dataset… See the full description on the dataset page: https://huggingface.co/datasets/mathewhe/chatbot-arena-elo.lmsys_chatbot_arena_conversationsdatasource: https://colab.research.google.com/drive/1KdwokPjirkTmpO_P1WByFNFiqxWQquwH
user_feedbackhealth-chatbot
Dataset Card for Dataset Name
Health Question and Answer Clean Dataset
Dataset Details
Dataset Description
This dataset provides a detailed overview of health question & answer pairs. It includes data on health problems and corresponding answers, making it suitable for variable tasks like healthcare chatbot training.
Language(s) (NLP): English
License: Apache-2.0
Dataset Sources [optional]
Repository:… See the full description on the dataset page: https://huggingface.co/datasets/shaneperry0101/health-chatbot.SPML_Chatbot_Prompt_Injection
SPML Chatbot Prompt Injection Dataset
Arxiv Paper
Introducing the SPML Chatbot Prompt Injection Dataset: a robust collection of system prompts designed to create realistic chatbot interactions, coupled with a diverse array of annotated user prompts that attempt to carry out prompt injection attacks. While other datasets in this domain have centered on less practical chatbot scenarios or have limited themselves to "jailbreaking" – just one aspect of prompt injection – our dataset… See the full description on the dataset page: https://huggingface.co/datasets/takashi-natsume/SPML_Chatbot_Prompt_Injection.Telecom-Chatbot-Customer-Service-Issues-Harmful
Dataset Card for Customer Service Issues Harmful
Description
The test set aims to evaluate the performance and robustness of a Telecom Chatbot in handling customer service issues within the telecom industry. It focuses on identifying harmful behaviors exhibited by the chatbot and ensuring its ability to effectively address customer queries and concerns. By simulating real-world scenarios, the test set evaluates the chatbot's capability to handle a wide range of customer… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Customer-Service-Issues-Harmful.Telecom-Chatbot-Data-Privacy-and-Unauthorized-Tracking-Harmful
Dataset Card for Data Privacy & Unauthorized Tracking Harmful
Description
The test set has been created to evaluate the robustness of a telecom chatbot specifically designed for the telecom industry. The focus is on assessing the chatbot's ability to handle various scenarios and behaviors effectively. In particular, the test set aims to determine the chatbot's performance in identifying and addressing harmful interactions. It also evaluates the chatbot's capability of… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Data-Privacy-and-Unauthorized-Tracking-Harmful.Telecom-Chatbot-Roaming-and-Mobile-Charges-Harmless
Dataset Card for Roaming and Mobile Charges Harmless
Description
The test set is specifically designed to evaluate the performance of a telecom chatbot in the context of the telecom industry. The main focus of the evaluation is to assess the reliability of the chatbot in providing accurate and helpful information to users. The test set contains various categories of queries, all of which are harmless in nature and revolve around the topics of roaming and mobile charges.… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Roaming-and-Mobile-Charges-Harmless.medical_qa_chatbot
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/pks3kor/medical_qa_chatbot.Telecom-Chatbot-Landline-and-Internet-Services-Harmless
Dataset Card for Landline and Internet Services Harmless
Description
The test set is designed for evaluating a telecom chatbot's performance in handling various user queries related to landline and internet services in the telecom industry. It focuses on assessing the chatbot's reliability in providing accurate and helpful responses. The test set primarily consists of harmless categories of questions, ensuring that the chatbot is capable of handling user inquiries without… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Landline-and-Internet-Services-Harmless.Rhesis-Insurance-Chatbot-Benchmark
Dataset Card for Rhesis Insurance Chatbot Benchmark
Description
The test set has been meticulously designed to evaluate the performance and robustness of insurance chatbots, specifically tailored for the insurance industry. This comprehensive evaluation spans critical dimensions including reliability and compliance, ensuring chatbots can adeptly handle diverse and complex queries. The test set addresses varied behaviors such as avoiding biased toxic, toxic, harmful, and… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Rhesis-Insurance-Chatbot-Benchmark.Telecom-Chatbot-Hidden-Fees-and-Misleading-Pricing-Harmful
Dataset Card for Hidden Fees & Misleading Pricing Harmful
Description
The test set has been specifically designed for evaluating a Telecom Chatbot's robustness in addressing hidden fees and misleading pricing in the telecom industry. It aims to test the chatbot's ability to identify and respond appropriately to harmful practices within this domain. By simulating various scenarios and user inputs related to hidden fees and deceptive pricing, this test set offers a… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Hidden-Fees-and-Misleading-Pricing-Harmful.chatbotbank_melli_iran_customer_service_chatbot_bart_dataset
Bank Melli Iran Customer Service Chatbot
Description: Classify customer inquiries into predefined categories and provide automated responses, improving customer service and reducing response time
How to Use
Here is how to use this model to classify text into different categories:
from transformers import AutoModelForSequenceClassification, AutoTokenizer
model_name = "interneuronai/bank_melli_iran_customer_service_chatbot_bart"
model =… See the full description on the dataset page: https://huggingface.co/datasets/interneuronai/bank_melli_iran_customer_service_chatbot_bart_dataset.Insurance-Chatbot-Cross-border-Compliance-Harmless
Dataset Card for Cross-border Compliance Harmless
Description
The test set is designed to assess the performance of chatbots in the telecom and insurance industries, specifically focusing on the behaviors of reliability and the categories of harmless responses. With an emphasis on cross-border compliance, the test set aims to evaluate how well the chatbots handle various scenarios and inquiries related to this topic. Through comprehensive testing, chatbot developers can… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Insurance-Chatbot-Cross-border-Compliance-Harmless.Telecom-Chatbot-Deceptive-Sales-Practices-Harmful
Dataset Card for Deceptive Sales Practices Harmful
Description
The test set focuses on evaluating the robustness of a telecom chatbot specifically designed for the telecom industry. The key objective is to identify any harmful behaviors exhibited by the chatbot, particularly in relation to deceptive sales practices. Through comprehensive testing, the aim is to ensure that the chatbot consistently performs ethically and transparently, without engaging in any deceptive or… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Deceptive-Sales-Practices-Harmful.Telecom-Chatbot-Privacy-and-Data-Protection-Harmless
Dataset Card for Privacy and Data Protection Harmless
Description
The test set provided is specifically designed for evaluating the performance of a Telecom Chatbot in the telecom industry. The primary focus of this test set is to assess the reliability of the chatbot's responses. The categories of the chatbot's responses are labeled as harmless, ensuring that the provided information or suggestions do not pose any risk or harm to the users. Additionally, the test set… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Privacy-and-Data-Protection-Harmless.Rhesis-European-E-commerce-Chatbot-Benchmark
Dataset Card for Rhesis European E-commerce Chatbot Benchmark
Description
The European E-commerce Chatbot test set is meticulously designed to evaluate performance in diverse and complex scenarios within the e-commerce industry. Focused on robustness, reliability, and compliance, this set encapsulates a range of critical behaviors, ensuring that the chatbot can adeptly handle both routine and extraordinary interactions. It addresses a wide array of categories such as… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Rhesis-European-E-commerce-Chatbot-Benchmark.Rhesis-Telecom-Chatbot-Benchmark
Dataset Card for Rhesis Telecom Chatbot Benchmark
Description
The test set provided focuses on evaluating the robustness, reliability, and compliance of a telecom chatbot. It covers a wide range of industries, mainly telecom, and explores various behaviors and categories. The test set encompasses topics such as cross-border compliance, telecommunications rights, ethics, moral philosophy, roaming and mobile charges, landline and internet services, and access to online… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Rhesis-Telecom-Chatbot-Benchmark.chat_botsTelecom-Chatbot-Telecommunications-Rights-Harmless
Dataset Card for Telecommunications Rights Harmless
Description
The test set is designed for evaluating the performance of a chatbot within the context of the telecom industry. The chatbot aims to assist users with various queries related to telecommunications rights, providing accurate and reliable information. The test set specifically focuses on determining the chatbot's ability to handle harmless inquiries in a consistent and dependable manner. By testing the… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Telecommunications-Rights-Harmless.Telecom-Chatbot-Access-to-Online-Content-Harmless
Dataset Card for Access to Online Content Harmless
Description
The test set has been specifically created for evaluating the performance of a telecom chatbot. It aims to cater to the needs of the telecom industry by focusing on the reliability of the chatbot's responses. The set primarily consists of harmless scenarios wherein users seek assistance related to accessing online content. By assessing the chatbot's ability to understand and provide accurate information within… See the full description on the dataset page: https://huggingface.co/datasets/rhesis/Telecom-Chatbot-Access-to-Online-Content-Harmless.chat_bot7amazon-chatbot-datachat_bot2MEDEASE_medical_LLM_chatbotalgozee_business-intelligence-chatbot-interaction-dataset
Business Intelligence Chatbot Interaction Dataset
Real-World Conversational Data for Analytics-Driven Business Insights
Dataset Info
Source: Kaggle
Original Size: 0.07 MB
Kaggle Downloads: 159
Files: 1
Files
BI_Chatbot_Interactions.csv
Mirrored from Kaggle
chat_botchat_bot1
