CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sadeem-ai /arabic-qna Sadeem QnA: An Arabic QnA Dataset 🌍✨ Welcome to the Sadeem QnA dataset, a vibrant collection designed for the advancement of Arabic natural language processing, specifically tailored for Question Answering (QnA) systems. Sourced from the rich and diverse content of Arabic Wikipedia, this dataset is a gateway to exploring the depths of Arabic language understanding, offering a unique challenge to both researchers and AI enthusiasts alike. About Sadeem QnA The Sadeem… See the full description on the dataset page: https://huggingface.co/datasets/sadeem-ai/arabic-qna.textquestion-answering1K<n<10K4 likes372 downloads3y agoHugging Face02arjunth2001 /online_privacy_qnaOnline Privacy Policy QnA Dataset textn<1K5 likes131 downloads5y agoHugging Face03neifuisan /Neuro-sama-QnAThis dataset was manually created, line by line, by my tiny hand! Why? Because I was just bored during my summer. textquestion-answeringn<1K50 likes116 downloads2y agoHugging Face04msamg /QnA_Descriptivetextn<1K0 likes103 downloads2y agoHugging Face05Qnancy /magicmotion MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance Quanhao Li*, Zhen Xing*, Rui Wang, Hui Zhang, Qi Dai, and Zuxuan Wu * equal contribution 💡 Abstract Recent advances in video generation have led to remarkable improvements in visual quality and temporal coherence. Upon this, trajectory-controllable video generation has emerged to enable precise object motion control through explicitly defined spatial paths. However, existing methods… See the full description on the dataset page: https://huggingface.co/datasets/Qnancy/magicmotion.imagen<1K0 likes89 downloads8mo agoHugging Face06agufsamudra /alodokter-qna Dataset Question Answer Health Indonesian Dataset Summary The Question Answer Health Indonesian dataset contains +250,000 question-and-answer pairs related to health topics sourced from the Alodokter website. The dataset spans a collection period from July 2023 to September 2023 (approximately 2 months). It is designed to facilitate research and development in the fields of natural language processing (NLP), particularly for Indonesian language models, health information… See the full description on the dataset page: https://huggingface.co/datasets/agufsamudra/alodokter-qna.text100K<n<1M4 likes78 downloads2y agoHugging Face07pgurazada1 /tesla-qna-feedback-logs Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/pgurazada1/tesla-qna-feedback-logs.textn<1K0 likes62 downloads2y agoHugging Face08atitaarora /qdrant_doc_qnatextn<1K1 likes60 downloads2y agoHugging Face09freococo /myanmar_qna_dataset Myanmar QnA Dataset v7 Language: Burmese (Myanmar)Total Entries: 22,783 QnA pairsTotal Sentences: ~ 466,330(Counted using the Myanmar sentence-ending symbol "။")License: CC0 1.0 (Public Domain) Description This dataset contains Myanmar-language question-answer pairs (QnA) generated with the assistance of ChatGPT-5 for question crafting with English and Gemini 3.0 Pro for Myanmar QnA generation. It is intended for research, AI training, and educational purposes. Each entry… See the full description on the dataset page: https://huggingface.co/datasets/freococo/myanmar_qna_dataset.tabularquestion-answering10K<n<100K0 likes45 downloads9mo agoHugging Face10abid /indonesia-medical-qna indonesia-medical-qna This dataset is configured so the Hugging Face Dataset Viewer loads qna.csv as the primary data file. The repository also contains all_links.csv, which has a different schema and is kept as an auxiliary file rather than part of the default viewer configuration. This avoids schema-casting errors caused by the Hub trying to combine both CSV files into one split. text100K<n<1M5 likes38 downloads6mo agoHugging Face11CoAILab /qna-quantum-information Q&A Quantum Information Dataset This dataset is created by digesting 500 different papers from the quantum information directory on arXiv, the papers are located based on their relevance to the quantum information keyword. Data retrieval The data is extracted from the .pdf files using PyMuPDF package proxied from langchain. Then the Q&A pair is generated by: Generate N questions per page of the .pdf document based on its content. We will feed each question to an LLM… See the full description on the dataset page: https://huggingface.co/datasets/CoAILab/qna-quantum-information.text10K<n<100K1 likes36 downloads2y agoHugging Face12bipinsaha /bangla-law-qnatextn<1K2 likes29 downloads2y agoHugging Face13coegggg /alodokter-qna Dataset Question Answer Health Indonesian Dataset Summary The Question Answer Health Indonesian dataset contains +250,000 question-and-answer pairs related to health topics sourced from the Alodokter website. The dataset spans a collection period from July 2023 to September 2023 (approximately 2 months). It is designed to facilitate research and development in the fields of natural language processing (NLP), particularly for Indonesian language models, health… See the full description on the dataset page: https://huggingface.co/datasets/coegggg/alodokter-qna.text100K<n<1M0 likes28 downloads4d agoHugging Face14AbhishekG13 /BNF_QNAtext10K<n<100K0 likes21 downloads2y agoHugging Face15Sumsam /QnA_for_Non-Technical_Roles Columns: Non-Technical Role: Specifies the role being assessed (e.g., Content Developer). Assessment Domain: Denotes the skill or attribute being evaluated (e.g., Adaptability). Question: The actual assessment question. Content Overview: The dataset is focused on assessing various competencies and skills relevant to non-technical roles. Questions are tailored to evaluate how individuals in these roles handle various situations and challenges. Example Entries: Role: Content Developer… See the full description on the dataset page: https://huggingface.co/datasets/Sumsam/QnA_for_Non-Technical_Roles.text1K<n<10K0 likes20 downloads3y agoHugging Face16AshtonLKY /workshop_QnAtextn<1K0 likes20 downloads3y agoHugging Face17vitoghif /wikipedia-id-qna Wikipedia ID Synthetic QnA This dataset contains synthetic question-answer pairs (QnA) generated using DeepSeek from Indonesian Wikipedia articles. The data has been sourced from this Wikipedia dataset, which contains a subset of Indonesian Wikipedia articles. Each entry includes a context, a related question and answer pair, and an unrelated question. Dataset Structure The dataset contains the following columns: id: A unique identifier for each row. context: A… See the full description on the dataset page: https://huggingface.co/datasets/vitoghif/wikipedia-id-qna.textquestion-answeringn<1K1 likes19 downloads2y agoHugging Face18rajveer43 /QnAMedicDaatasetgatedtextquestion-answering10K<n<100K1 likes18 downloads2y agoHugging Face19atitaarora /qdrant_docs_qna_ragastextn<1K0 likes18 downloads3y agoHugging Face20SamagraDataGov /QnA_20240513_122841textn<1K0 likes18 downloads2y agoHugging Face21bhatthars /nbme_qnaquestion-context-answer format rows for nbme qna fine-tuning text1K<n<10K0 likes18 downloads2y agoHugging Face22Sid3503 /Human-Like-Gut-Health-DPO-QnA Gut Health DPO Dataset Overview This dataset contains 200 carefully curated examples for Direct Preference Optimization (DPO) training in the domain of gut health and digestive wellness. Each example consists of a user prompt, a "chosen" response (preferred), and a "rejected" response (less preferred), designed to train AI models to provide high-quality, medically responsible advice on digestive health topics. Dataset Structure The dataset is provided in CSV… See the full description on the dataset page: https://huggingface.co/datasets/Sid3503/Human-Like-Gut-Health-DPO-QnA.texttext-classificationn<1K2 likes18 downloads11mo agoHugging Face23kaifahmad /network-QnA-datasettabular1K<n<10K3 likes17 downloads3y agoHugging Face24abhishekdey /hr_assistant_qna_datasettextquestion-answeringn<1K0 likes17 downloads1y agoHugging Face25catalin1122 /wiki-ro-qna Description There are more than 550k questions with roughly 53k paragraphs. The questions were built using the ChatGPT 3.5 API. The dataset is based on the Romanian Wikipedia 2020 June dump, curated by Dumitrescu Stefan. The paragraphs retained are those between 100 and 410 words (roughly 512 max tokens), using the following script: # Open the text file with open('wiki-ro/corpus/wiki-ro/wiki.txt.train', 'r') as file: # Read the entire content… See the full description on the dataset page: https://huggingface.co/datasets/catalin1122/wiki-ro-qna.texttable-question-answering10K<n<100K2 likes16 downloads2y agoHugging Face26elvanalabs /customer-review-qna-25 📦 Dataset: customer-review-qna-25 🧾 Summary A small dataset of 25 GPT-4o-mini generated customer reviews, each containing: input_question: Prompt or user query output: Generated customer review context: Background info (e.g. product, tone) tags: Labels like positive, complaint, delivery, etc. ✅ safe ✅ Filtered ✅ Compact & usable for review generation or sentiment tasks 🧱 Dataset Structure Format: CSVSize: 25 samplesFields: Field Type… See the full description on the dataset page: https://huggingface.co/datasets/elvanalabs/customer-review-qna-25.texttext-generationn<1K2 likes16 downloads1y agoHugging Face27Adhelard /Neuro-sama-QnAThis dataset was manually created, line by line, by my tiny hand! Why? Because I was just bored during my summer. textquestion-answeringn<1K1 likes16 downloads7mo agoHugging Face28ksh-nyp /tcm-qnatextn<1K3 likes15 downloads3y agoHugging Face29CopyleftCultivars /KNF-Methods-QnAWork in progress. Not yet reviewed by domain experts. textquestion-answeringn<1K0 likes14 downloads2y agoHugging Face30Kliro /Neuro-sama-QnAThis dataset was manually created, line by line, by my tiny hand! Why? Because I was just bored during my summer. textquestion-answeringn<1K0 likes14 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.