CoolFace
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01JSALT2026-Conv-AI-Simulator /turnbench-dev-no-backchannel TurnBench Dev - Backchannels Removed A derivative of mundo-ai/turn-benchmark-dev with every majority-annotated backchannel removed from the audio: 1853 backchannels across 38 conversations, 2077.0 seconds in total, cut out of the speaker's own channel and replaced by background noise taken from elsewhere in that same channel. Everything else is the original recording, sample for sample. Same conversations, same duration, same timeline, same annotator tracks, same speech -- only… See the full description on the dataset page: https://huggingface.co/datasets/JSALT2026-Conv-AI-Simulator/turnbench-dev-no-backchannel.audiovoice-activity-detectionn<1K0 likes112 downloads12d agoHugging Face02PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-Shuffledtextn<1K0 likes17 downloads2y agoHugging Face03PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguard-splittext1K<n<10K0 likes15 downloads1y agoHugging Face04PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite You should mask everything except the last turn. The only part that matters to teach the model is the last turn, as you are teaching it to always output thinking, no matter what the user feeds it. It's setup to be trained like R1: text1K<n<10K1 likes14 downloads2y agoHugging Face05mlfoundations-dev /dedup_ablation_sim_threshold_0tabular10K<n<100K0 likes11 downloads2y agoHugging Face06TAUR-dev /D-EVAL__standard_eval_v3__simple_test__exp_runner_2-eval_0 D-EVAL__standard_eval_v3__simple_test__exp_runner_2-eval_0 This evaluation dataset was created as part of the simple_test__exp_runner_2 experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "Qwen/Qwen2.5-1.5b-Instruct", "tasks": ["countdown_2arg", "countdown_3arg"], "annotators": ["greedy"], "splits": ["test"], "dataset_url":… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__simple_test__exp_runner_2-eval_0.textn<1K0 likes9 downloads1y agoHugging Face07PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-Shuffledtext1K<n<10K0 likes8 downloads5mo agoHugging Face08PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all You should mask everything except the last turn. The only part that matters to teach the model is the last turn, as you are teaching it to always output thinking, no matter what the user feeds it. Generation Script This is what I used to generate the dataset, so that it's setup to be trained like R1: import requests import json import time import pandas as pd from datasets import load_dataset from tqdm import… See the full description on the dataset page: https://huggingface.co/datasets/PJMixers-Dev/lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all.textn<1K0 likes8 downloads2y agoHugging Face09TAUR-dev /D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft This evaluation dataset was created as part of the simple_test__exp_runner_3 experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/M-simple_test__exp_runner_3-sft", "tasks": ["countdown_2arg", "countdown_3arg"], "annotators": ["greedy"], "splits": ["test"]… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft.textn<1K0 likes8 downloads1y agoHugging Face10azain /LibriTTS-dev-clean-16khz-mono-loudnorm-100-random-samples-2024-04-18-17-34-39-similaritiestext1K<n<10K0 likes6 downloads2y agoHugging Face11mlfoundations-dev /dedup_ablation_sim_threshold_10tabular10K<n<100K0 likes6 downloads2y agoHugging Face12mlfoundations-dev /dedup_ablation_sim_threshold_20tabular10K<n<100K0 likes6 downloads2y agoHugging Face13mlfoundations-dev /dedup_ablation_sim_threshold_40tabular10K<n<100K0 likes6 downloads2y agoHugging Face14mlfoundations-dev /dedup_ablation_sim_threshold_nonetabular10K<n<100K0 likes5 downloads2y agoHugging Face15PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguard-split-qwqtabularn<1K0 likes4 downloads1y agoHugging Face16TAUR-dev /D-SFT_C-back_to_og_mix__simple_retries__sbon-sft-datatext1K<n<10K0 likes4 downloads1y agoHugging Face17TAUR-dev /D-EVAL__standard_eval_v3__back_to_og_mix__simple_mix__rl_eval-eval_rl D-EVAL__standard_eval_v3__back_to_og_mix__simple_mix__rl_eval-eval_rl This evaluation dataset was created as part of the back_to_og_mix__simple_mix__rl_eval experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/SIE-back_to_og_mix__simple_retries__sbon-rl", "tasks": ["countdown_2arg", "countdown_3arg", "countdown_4arg"… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__back_to_og_mix__simple_mix__rl_eval-eval_rl.text1K<n<10K0 likes4 downloads1y agoHugging Face18TAUR-dev /dataset__acronym_generation__simple__6_wordstext1K<n<10K0 likes4 downloads1y agoHugging Face19TAUR-dev /D-EVAL__simple_eval__cd3arg-sft1ep_grpo_1e6lr_mix_ss_pse_vote_ansrev-sfttext1K<n<10K0 likes3 downloads1y agoHugging Face20TAUR-dev /D-simple_test-sft-datatextn<1K0 likes3 downloads1y agoHugging Face21TAUR-dev /D-back_to_og_mix__simple_retries__sbon-sft-datatext1K<n<10K0 likes3 downloads1y agoHugging Face22TAUR-dev /dataset__acronym_generation__simple__4_wordstext1K<n<10K0 likes3 downloads1y agoHugging Face23TAUR-dev /dataset__acronym_generation__simple__5_wordstext1K<n<10K0 likes3 downloads1y agoHugging Face24TAUR-dev /dataset__acronym_generation__simple__7_wordstext1K<n<10K0 likes3 downloads1y agoHugging Face25PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite-filttext1K<n<10K0 likes2 downloads2y agoHugging Face26TAUR-dev /D-EVAL__standard_eval_v3__back_to_og_mix__simple_retries__sbon-eval_sft D-EVAL__standard_eval_v3__back_to_og_mix__simple_retries__sbon-eval_sft This evaluation dataset was created as part of the back_to_og_mix__simple_retries__sbon experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/M-back_to_og_mix__simple_retries__sbon-sft", "tasks": ["countdown_2arg", "countdown_3arg", "countdown_4arg"… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__back_to_og_mix__simple_retries__sbon-eval_sft.text1K<n<10K0 likes2 downloads1y agoHugging Face27PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite-classifiedgatedtext1K<n<10K0 likes1 downloads1y agoHugging Face28PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguardgatedtext10K<n<100K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.