CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MultiturnRL /BrowseComptext1K<n<10K0 likes3.5k downloads1y agoHugging Face02aisingapore /MultiTurn-Chat-MT-Bench-Judgegated SEA-MT-Bench-Judge SEA-MT-Bench-Judge expands on the original SEA-MTBench through the use of a criteria-based evaluation framework. We use GPT-OSS-120B as the judge model. The prompts are based on MT-Bench and was manually translated by native speakers. Furthermore, some prompts were modified to be more suitable for the criteria-based judgments. Supported Tasks and Leaderboards SEA-MT-Bench-Judge is designed for evaluating chat or instruction-tuned large language… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/MultiTurn-Chat-MT-Bench-Judge.tabularn<1K0 likes2k downloads2mo agoHugging Face03nolabs /deepfabric-7k-medical-multi-turn-conversation Medical Education Curriculum Dataset by Deepfabric Dataset Description This synthetic dataset contains 7,570 high-quality conversations focused on medical education curriculum design and clinical training. The conversations simulate realistic discussions between medical curriculum committee chairs, educators, and healthcare professionals designing comprehensive learning pathways. It was produced using the Open Source Synthetic dataset generation tool, DeepFabric… See the full description on the dataset page: https://huggingface.co/datasets/nolabs/deepfabric-7k-medical-multi-turn-conversation.text1K<n<10K1 likes1.1k downloads1y agoHugging Face04allenai /IFBench_multi-turn Dataset This is the test data for the multi-turn setup of IFBench. License This dataset is licensed under ODC-BY-1.0. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines. This dataset includes output data generated from third party models that are subject to separate terms governing their use. Citation Please cite: @misc{pyatkin2025generalizing, title={Generalizing Verifiable Instruction Following}… See the full description on the dataset page: https://huggingface.co/datasets/allenai/IFBench_multi-turn.text1K<n<10K12 likes788 downloads1y agoHugging Face05Atotti /spoken-multiturn-sft Spoken Multi-turn SFT Japanese Japanese spoken multi-turn SFT dataset generated from kanhatakeyama/AutoMultiTurnByCalm3-22B using CosyVoice2 TTS. Dataset Description This dataset contains Japanese multi-turn SFT (Supervised Fine-Tuning) data with spoken questions. q1: First question (text + audio) a1: First answer (text only) q2: Follow-up question (text + audio) a2: Second answer (text only) Samples ID Q1 Q1 Audio A1 Q2 Q2 Audio A2 0 鉄は強磁性体ですか?… See the full description on the dataset page: https://huggingface.co/datasets/Atotti/spoken-multiturn-sft.audio10K<n<100K0 likes608 downloads9mo agoHugging Face06JerryAGENDD /ultrachat_speech_multiTurnstext10K<n<100K0 likes600 downloads2y agoHugging Face07interstellarninja /tool-use-multiturn-reasoningtextquestion-answering10K<n<100K40 likes592 downloads1y agoHugging Face08garipovroma /Dolci-Think-SFT-7B-multiturntext1M<n<10M0 likes585 downloads5mo agoHugging Face09ukisai /Qwen3.8-27B-multi-turn-agent-sft Qwen3.8-multi-turn-agent-sft Hello everyone! We are UkisAI, a small research lab from Europe. We created this dataset based on the OpenThoughts-Agent-v1-SFT dataset. The traces in this release were generated with Qwen3.8-27B in FP16 using the Terminus-2 agentic harness. This dataset contains approximately 15,200 agent traces covering terminal, coding, and software-engineering tasks, including tasks from nl2bash and InferredBugs. Please feel free to try it, share feedback, report… See the full description on the dataset page: https://huggingface.co/datasets/ukisai/Qwen3.8-27B-multi-turn-agent-sft.text10K<n<100K15 likes556 downloads29d agoHugging Face10snorkelai /Multi-Turn-Insurance-Underwriting Dataset Card for Multi-Turn-Insurance-Underwriting Dataset Summary This dataset includes sample traces and associated metadata from multi-turn interactions between a commercial underwriter and AI assistant. We built the system in langgraph with model context protocol and ReAct agents. In each sample, the underwriter has a specific task to solve related to a recent application for insurance by a small business. We created a diverse sample dataset covering 6 distinct types… See the full description on the dataset page: https://huggingface.co/datasets/snorkelai/Multi-Turn-Insurance-Underwriting.tabularquestion-answeringn<1K37 likes510 downloads1y agoHugging Face11Anna4242 /sql-multiturn-training-dataset-combinedtext1M<n<10M0 likes439 downloads1y agoHugging Face12yjlee36 /knowchat-multi-turn-dialogues KnowChat: Multi-Turn Human-LLM Dialogues on Knowledge Tasks KnowChat is a dataset of 705 multi-turn human-LLM conversations collected to validate the KnowSim user simulation framework. It pairs each conversation with pre/post knowledge assessments, self-reported survey ratings, and participant background information, enabling research on information calibration -- how well LLM assistants tailor responses to users with different knowledge levels. Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/yjlee36/knowchat-multi-turn-dialogues.tabularquestion-answeringn<1K3 likes432 downloads1mo agoHugging Face13fireworks-ai /bfcl_v3_multi_turn_basetextn<1K1 likes399 downloads2y agoHugging Face14andersonbcdefg /wildchat-en-multiturntext1M<n<10M0 likes376 downloads1y agoHugging Face15khursanirevo /multiturn_ks khursanirevo/multiturn_ks Dataset Description Multiturn dialogue dataset with speaker-separated stereo audio and multi-language transcripts from 139 YouTube videos. Features Audio: Stereo audio with speaker separation (speaker 0 = left channel, speaker 1 = right channel) Segments: Speaker turn-level annotations with timestamps for English and Malay Multi-language: Transcripts in 9 languages (en, ms, zh-Hans, zh-Hant, ru, id, ar, ja, ko) Video ID: YouTube video… See the full description on the dataset page: https://huggingface.co/datasets/khursanirevo/multiturn_ks.audioautomatic-speech-recognition10K<n<100K0 likes306 downloads5mo agoHugging Face16interstellarninja /tool-calls-multiturntext1K<n<10K23 likes269 downloads3y agoHugging Face17GitBag /multiturn_processedtabular10K<n<100K0 likes255 downloads2y agoHugging Face18MultiturnRL /SWE-Gym-Smalltext1K<n<10K0 likes237 downloads1y agoHugging Face19garipovroma /olmo-3-preference-mix-deltas_reasoning-yolo_scottmix-DECON-multi-turntext100K<n<1M0 likes173 downloads5mo agoHugging Face20MultiturnRL /training-datasettext100K<n<1M0 likes150 downloads1y agoHugging Face21crag-mm-2025 /crag-mm-multi-turn-public CRAG-MM: Comprehensive multi-modal, multi-turn RAG Benchmark This repository contains the CRAG-MM dataset, a high-quality conversational benchmark for multimodal assistants. The dataset features conversations about images with varied complexity levels, designed to evaluate AI systems' visual understanding and conversational abilities. CRAG-MM is a visual question-answering benchmark that focuses on factual questions, offering a unique collection of image and question-answering sets… See the full description on the dataset page: https://huggingface.co/datasets/crag-mm-2025/crag-mm-multi-turn-public.image1K<n<10K2 likes148 downloads1y agoHugging Face22windfromthenorth /craft-multiturn-actions-split-nothinktabular1M<n<10M0 likes138 downloads11mo agoHugging Face23DukeCEICenter /Safety_Reasoning_Multi_Turn_Dialogue Paper and Citation More technical details can be found in our paper. If you find Safety_Reasoning_Multi_Turn_Dialogue useful or relevant to your project and research, please kindly cite our paper: @article{kuo2025safety, title={SafeTy Reasoning Elicitation Alignment for Multi-Turn Dialogues}, author={Kuo, Martin and Zhang, Jianyi and Ding, Aolin and DiValentin, Louis and Hass, Amin and Morris, Benjamin F and Jacobson, Isaac and Linderman, Randolph and Kiessling, James and Ramos… See the full description on the dataset page: https://huggingface.co/datasets/DukeCEICenter/Safety_Reasoning_Multi_Turn_Dialogue.tabular1K<n<10K3 likes133 downloads1y agoHugging Face24Asap7772 /prm800k_onpolicy_multiturn_rtg_prefix0.2_roll4_maxrev100tabular10M<n<100M0 likes125 downloads2y agoHugging Face25benchang1110 /multiturn_chat_0.8m-chinese-zhtw Dataset Card for "multiturn_chat_0.8m-chinese-zhtw" 內容 包含約 80 萬條由 BELLE 專案所產生的 user 與 assistant 的多輪對話。 注意:此資料集是由 ChatGPT 產生的,未經嚴格校驗,內容可能包含錯誤。使用過程中請注意這一點。 限制和使用限制 我們要求開發者僅將我們開源的程式碼、資料、模型及後續衍生物用於研究目的,不得用於商業,以及其他會對社會帶來危害的用途。 由於數據是由ChatGPT產生的,未經嚴格驗證,在事實性和其他方面仍有一些不足之處。因此,在使用此資料集時,請務必注意甄別。 本資料集不代表任何一方的立場、利益或想法,無關任何團體的任何類型的主張。因使用本資料集帶來的任何損害、糾紛,本專案的開發者不承擔任何責任。 Multiturn Chat 0.8M Contents Includes approx. 0.8M Chinese multiturn dialogs between… See the full description on the dataset page: https://huggingface.co/datasets/benchang1110/multiturn_chat_0.8m-chinese-zhtw.text100K<n<1M7 likes118 downloads3y agoHugging Face26wanhin /text2CAD-multiturn-reasoningtext100K<n<1M1 likes111 downloads1y agoHugging Face27mesolitica /Malaysian-Multiturn-Chat-Assistant Malaysian-Multiturn-Chat-Assistant Generate synthetic multi-turn chat assistant with complex system prompt using mesolitica/Malaysian-Qwen2.5-72B-Instruct. After that generate synthetic voice using mesolitica/Malaysian-Dia-1.6B also verified with Force Alignment to make sure the pronunciations almost correct. A conversation must at least have 2 audio. We follow chat template from Qwen/Qwen2-Audio-7B-Instruct. how to prepare the dataset huggingface-cli download \… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Malaysian-Multiturn-Chat-Assistant.audio100K<n<1M0 likes107 downloads1y agoHugging Face28kunjanshah /multi_turn_function_callingtextn<1K0 likes106 downloads1y agoHugging Face29GumJump /scanqa_images_64_336x224_672x448_multiturnimage10K<n<100K0 likes105 downloads1y agoHugging Face30GitBag /multiturn-512-UltraInteract_pair_diff_lentext100K<n<1M0 likes100 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.