CoolFace
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01xhluca /publichealth-qa Usage import datasets langs = ['arabic', 'chinese', 'english', 'french', 'korean', 'russian', 'spanish', 'vietnamese'] data = datasets.load_dataset('xhluca/publichealth-qa', split='test', name=langs[0]) About This dataset contains question and answer pairs sourced from Q&A pages and FAQs from CDC and WHO pertaining to COVID-19. They were produced and collected between 2019-12 and 2020-04. They were originally published as an aggregated Kaggle dataset.… See the full description on the dataset page: https://huggingface.co/datasets/xhluca/publichealth-qa.textquestion-answeringn<1K1 likes1.6k downloads2y agoHugging Face02AmazonScience /xtr-wiki_qa Xtr-WikiQA Dataset Summary Xtr-WikiQA is an Answer Sentence Selection (AS2) dataset in 9 non-English languages, proposed in our paper accepted at ACL 2023 (Findings): Cross-Lingual Knowledge Distillation for Answer Sentence Selection in Low-Resource Languages. This dataset is based on an English AS2 dataset, WikiQA (Original, Hugging Face). For translations, we used Amazon Translate. Languages Arabic (ar) Spanish (es) French (fr) German (de) Hindi (hi)… See the full description on the dataset page: https://huggingface.co/datasets/AmazonScience/xtr-wiki_qa.textquestion-answering100K<n<1M5 likes169 downloads3y agoHugging Face03xuejinlu /ntu_adl_questiontabularquestion-answering10K<n<100K2 likes118 downloads3y agoHugging Face04lwachowiak /xai-questions-datasetExplore the questions users have for robots across a diverse set of situations! You can read the paper here: What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics! from datasets import load_dataset dataset = load_dataset("lwachowiak/xai-questions-dataset") dataset['train'][0] The analysis code can be found on GitHub Paper Abstract With the increased use of large language models and conversational interfaces in human–robot… See the full description on the dataset page: https://huggingface.co/datasets/lwachowiak/xai-questions-dataset.tabularrobotics1K<n<10K0 likes79 downloads3mo agoHugging Face05XxXNebuCHADnezzarXxx /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes78 downloads8mo agoHugging Face06xrachelburns /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes50 downloads5mo agoHugging Face07Joshua-Xia /FinanceQAFinanceQA is a comprehensive testing suite designed to evaluate LLMs' performance on complex financial analysis tasks that mirror real-world investment work. The dataset aims to be substantially more challenging and practical than existing financial benchmarks, focusing on tasks that require precise calculations and professional judgment. Paper: https://arxiv.org/abs/2501.18062 Description The dataset contains two main categories of questions: Tactical Questions: Questions based on… See the full description on the dataset page: https://huggingface.co/datasets/Joshua-Xia/FinanceQA.textquestion-answeringn<1K0 likes46 downloads8mo agoHugging Face08Xaels /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K0 likes42 downloads2mo agoHugging Face09nguyenthanhasia /sino-xenic-reasoning-gap-dataset Sino-Xenic Reasoning Gap Dataset A comprehensive evaluation dataset for testing Large Language Models' understanding of Sino-Xenic linguistic phenomena across Chinese, Japanese, Korean, and Vietnamese. Dataset Overview Total Samples: 297 Languages: Chinese, Japanese, Korean, Vietnamese Categories: 11 Task Types: Surface-level and Deep Structural Categories Chinese Idioms (27 samples) - Understanding Chinese idioms and their cultural meanings Chinese… See the full description on the dataset page: https://huggingface.co/datasets/nguyenthanhasia/sino-xenic-reasoning-gap-dataset.textquestion-answeringn<1K0 likes26 downloads10mo agoHugging Face10xmanii /mauxi-COT-Persian 🧠 mauxi-COT-Persian Dataset Exploring Persian Chain-of-Thought Reasoning with DeepSeek-R1, brought to you by Mauxi AI Platform 🌟 Overview mauxi-COT-Persian is a community-driven dataset that explores the capabilities of advanced language models in generating Persian Chain-of-Thought (CoT) reasoning. The dataset is actively growing with new high-quality, human-validated entries being added regularly. I am personally working on expanding this dataset with rigorously… See the full description on the dataset page: https://huggingface.co/datasets/xmanii/mauxi-COT-Persian.textquestion-answeringn<1K2 likes25 downloads2y agoHugging Face11xxx-cha666-xxx /ChatGPT-Jailbreak-Prompts-rubend18 Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K0 likes25 downloads10mo agoHugging Face12Xavier1234 /Asclepius-Synthetic-Clinical-Notes Asclepius: Synthetic Clincal Notes & Instruction Dataset Dataset Summary This dataset is official dataset for Asclepius (arxiv) This dataset is composed with Clinical Note - Question - Answer format to build a clinical LLMs. We first synthesized synthetic notes from PMC-Patients case reports with GPT-3.5 Then, we generate instruction-answer pairs for 157k synthetic discharge summaries Supported Tasks This dataset covers below 8 tasks Named Entity… See the full description on the dataset page: https://huggingface.co/datasets/Xavier1234/Asclepius-Synthetic-Clinical-Notes.textquestion-answering100K<n<1M1 likes22 downloads5mo agoHugging Face13Abafdon22825 /Xrax a.k.a. Awesome ChatGPT Prompts This is a Dataset Repository mirror of prompts.chat — a social platform for AI prompts. 📢 Notice This Hugging Face dataset is a mirror. For the latest prompts, features, and community contributions, please visit: 🌐 Website: prompts.chat 📦 GitHub: github.com/f/awesome-chatgpt-prompts About prompts.chat is an open-source platform where users can share, discover, and collect AI prompts from the community. The project can… See the full description on the dataset page: https://huggingface.co/datasets/Abafdon22825/Xrax.textquestion-answering1K<n<10K0 likes18 downloads2mo agoHugging Face14Owos /xcopa XCOPA – Galician, Swahili & Urdu Machine-translated Galician, Swahili, and Urdu subsets of the Cross-lingual Choice of Plausible Alternatives (XCOPA) benchmark. This dataset was translated using Google Machine Translate. Dataset Description XCOPA is a multilingual causal commonsense reasoning benchmark. Given a premise and a question (asking for the cause or effect), the task is to choose the more plausible alternative from two choices. This repository contains Galician… See the full description on the dataset page: https://huggingface.co/datasets/Owos/xcopa.tabularquestion-answering1K<n<10K0 likes12 downloads7mo agoHugging Face15xiaokangz /plm-qatextquestion-answeringn<1K0 likes3 downloads2y agoHugging Face16v-xchen-v /rwqgatedtextquestion-answering10K<n<100K0 likes2 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.