CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01baptistefrancois1 /s2s-fr-finetuning s2s-fr-finetuning Corpus FR pour le finetuning speech-to-speech (Liquid-Audio / LFM2-Audio), construit par une pipeline de prétraitement : VAD, ASR + alignement mot, segmentation aux frontières de mots, filtrage qualité perceptuelle, normalisation de texte, déduplication. Utilisation from datasets import load_dataset ds = load_dataset("baptistefrancois1/s2s-fr-finetuning", "common_voice_fr") Un config HF par source d'origine : common_voice_fr, emilia_yodas_fr… See the full description on the dataset page: https://huggingface.co/datasets/baptistefrancois1/s2s-fr-finetuning.audio100K<n<1M0 likes254 downloads1mo agoHugging Face02Maisum-Abbas-123 /Urdu-Finetuning-Data-VibeVoice-Largeaudio10K<n<100K0 likes179 downloads8mo agoHugging Face03laion /emotional-roleplay-finetuning-dataset Artificial Voice Roleplay Dataset 67,491 fully-synthetic speech clips (~184 hours) pairing expressive role-play / character voice-direction captions with generated audio, across German, English, Spanish, and French (German-dominant). Rich in exaggerated fantasy/creature voices (orc, goblin, troll, ogre, zombie, dragon, demon, witch, banshee, imp, fairy, gnome, robot, murloc, harpy, skeleton, ghost, vampire …) and high-arousal emotional delivery (rage, fear, grief, menace). Every… See the full description on the dataset page: https://huggingface.co/datasets/laion/emotional-roleplay-finetuning-dataset.audiotext-to-speech10K<n<100K4 likes122 downloads2mo agoHugging Face04MAdel121 /Common-Voice-17-Arabic-for-Seasme-CSM-Finetuning Curated Arabic Speech Dataset for Seasme (from MCV17) Dataset Description This dataset is a curated and preprocessed version of the Arabic (ar) subset from Mozilla Common Voice (MCV) 17.0. It has been specifically prepared for fine-tuning conversational speech models, with a primary focus on the Seasme-CSM model architecture. The dataset consists of audio clips in WAV format (24kHz, mono) and their corresponding transcripts, along with integer speaker IDs. The original… See the full description on the dataset page: https://huggingface.co/datasets/MAdel121/Common-Voice-17-Arabic-for-Seasme-CSM-Finetuning.audio10K<n<100K1 likes67 downloads1y agoHugging Face05demegire /personaplex-finetuning-pharma-data-sample PersonaPlex Finetuning — Pharma Data Sample A 10-example slice of the synthetic patient-support / medication adherence dataset used to train demegire/personaplex-finetune-pharma. The on-disk layout below is exactly what the trainer in emotion-machine-org/personaplex-finetune consumes — use this as a template when building your own. Split: 8 train / 2 eval (mirrors the upstream 2003 / 20 split at sample scale). Layout . ├── adhery_v2.jsonl # master… See the full description on the dataset page: https://huggingface.co/datasets/demegire/personaplex-finetuning-pharma-data-sample.audiotext-to-speechn<1K0 likes65 downloads4mo agoHugging Face06OpenWhistleNeurIPS26 /OpenWhistle-Classification-Finetuning OpenWhistle Classification Finetuning Dataset OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning is the public classification finetuning dataset used for dolphin whistle identity classification. It contains short whistle clips, whistle-level metadata, fundamental-frequency tracks, rendered F0 spectrograms, and integer class labels. The main reviewer-facing subset is the balanced balanced config. It contains six classes: NSW_1 (label=0) SW_Luna (label=1) SW_Nana (label=2)… See the full description on the dataset page: https://huggingface.co/datasets/OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning.audioaudio-classification10K<n<100K0 likes63 downloads5mo agoHugging Face07AhmedRezik /fineTuningDataaudio10K<n<100K0 likes33 downloads3y agoHugging Face08tonypeng /whisper-finetuningaudio1K<n<10K0 likes27 downloads2y agoHugging Face09Dev523 /tts-finetuningaudion<1K0 likes19 downloads1y agoHugging Face10B0808 /MDbA_FineTuningaudion<1K0 likes16 downloads3y agoHugging Face11dngngnguyencode /TTS_finetuning_datasetaudio10K<n<100K0 likes16 downloads2y agoHugging Face12thomaslu /articulationGAN_finetuning_data Dataset Card for "articulationGAN_finetuning_data" More Information needed audion<1K0 likes15 downloads3y agoHugging Face13classen3 /whisper-finetuning-for-aseeaudion<1K0 likes15 downloads3y agoHugging Face14B0808 /MDbA-FineTuningV2audion<1K0 likes12 downloads3y agoHugging Face15Kady-x /nova_finetuningaudion<1K0 likes12 downloads1y agoHugging Face16deboleen6 /whisper_finetuningaudion<1K0 likes11 downloads2y agoHugging Face17YongJaeLee /Whisper_FineTuning_Koaudio10K<n<100K0 likes10 downloads1y agoHugging Face18YongJaeLee /Whisper_FineTuning_Suaudio10K<n<100K0 likes10 downloads1y agoHugging Face19OleksandrAbashkin /fine-tuning-Jewish Dataset Card for "fine-tuning-Jewish" More Information needed audion<1K0 likes9 downloads2y agoHugging Face20wahyuachmad /Zeta-Voice-ID-Vits-Finetuningaudion<1K0 likes9 downloads2y agoHugging Face21sravan-gorugantu /techolution_noun_finetuning_v1audion<1K0 likes9 downloads2y agoHugging Face22anchaeyeon /whisper_finetuningaudio1K<n<10K0 likes9 downloads1y agoHugging Face23Shubham079 /whisper-finetuning-audio-filesaudio10K<n<100K0 likes8 downloads1y agoHugging Face24tonypeng /whisper-finetuning-testaudion<1K0 likes6 downloads2y agoHugging Face25LasseRogers2111 /stt_lowrank_finetuningaudion<1K0 likes6 downloads1y agoHugging Face26OpenWhistleNeurIPS26 /OpenWhistle-Detection-Finetuning OpenWhistle Detection Finetuning Expert-annotated whistle-type detection dataset for OpenWhistle. Each example is a fixed-length 0.5 s audio window labeled with the whistle types present in that window; background/no-whistle windows are represented by an all-zero target vector. Overview Task: multi-label whistle-type detection on fixed-length audio windows Target vector: label, with one binary decision per whistle type in the order SW_Neo, SW_Luna, SW_Nikita, SW_Nana… See the full description on the dataset page: https://huggingface.co/datasets/OpenWhistleNeurIPS26/OpenWhistle-Detection-Finetuning.audio1K<n<10K0 likes6 downloads5mo agoHugging Face27CarjGilson /fineTuningaudion<1K0 likes5 downloads2y agoHugging Face28Party-Lemur /Fine-Tuning-03audion<1K0 likes5 downloads9mo agoHugging Face29KisharXCoder /parler-ttts-dataset-for-finetuningaudio10K<n<100K0 likes5 downloads5mo agoHugging Face30dianpham /datasets_finetuning_PhoWhisperaudio1K<n<10K1 likes4 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.