datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Urdu-Finetuning-Data-VibeVoice-Largeemotional-roleplay-finetuning-dataset
Artificial Voice Roleplay Dataset
67,491 fully-synthetic speech clips (~184 hours) pairing expressive role-play / character
voice-direction captions with generated audio, across German, English, Spanish, and French
(German-dominant). Rich in exaggerated fantasy/creature voices (orc, goblin, troll, ogre,
zombie, dragon, demon, witch, banshee, imp, fairy, gnome, robot, murloc, harpy, skeleton, ghost,
vampire …) and high-arousal emotional delivery (rage, fear, grief, menace).
Every… See the full description on the dataset page: https://huggingface.co/datasets/laion/emotional-roleplay-finetuning-dataset.s2s-fr-finetuning
s2s-fr-finetuning
Corpus FR pour le finetuning speech-to-speech (Liquid-Audio / LFM2-Audio), construit par une
pipeline de prétraitement : VAD, ASR + alignement mot, segmentation aux frontières de mots,
filtrage qualité perceptuelle, normalisation de texte, déduplication.
Utilisation
from datasets import load_dataset
ds = load_dataset("baptistefrancois1/s2s-fr-finetuning", "common_voice_fr")
Un config HF par source d'origine : common_voice_fr, emilia_yodas_fr… See the full description on the dataset page: https://huggingface.co/datasets/baptistefrancois1/s2s-fr-finetuning.FeruzaSpeech_to_fine_tuning
FeruzaSpeech_to_fine_tuning
A speech corpus of ⏱️ ~59.1 total hours of Uzbek audio paired with Latin‑script transcripts, intended for fine‑tuning ASR / speech‑to‑text models.
Dataset Details
Dataset Description
This dataset contains recordings of native Uzbek speakers reading a mix of classical literature excerpts and school‑level writing prompts:
001: Choliqushi (a novel by Rashod Nuri Guntekin, trans. by Mirzakalon Ismoiliy; first pub. Sept 1900).
002:… See the full description on the dataset page: https://huggingface.co/datasets/nickoo004/FeruzaSpeech_to_fine_tuning.personaplex-finetuning-pharma-data-sample
PersonaPlex Finetuning — Pharma Data Sample
A 10-example slice of the synthetic patient-support / medication
adherence dataset used to train
demegire/personaplex-finetune-pharma.
The on-disk layout below is exactly what the trainer in
emotion-machine-org/personaplex-finetune
consumes — use this as a template when building your own.
Split: 8 train / 2 eval (mirrors the upstream 2003 / 20 split at
sample scale).
Layout
.
├── adhery_v2.jsonl # master… See the full description on the dataset page: https://huggingface.co/datasets/demegire/personaplex-finetuning-pharma-data-sample.Common-Voice-17-Arabic-for-Seasme-CSM-Finetuning
Curated Arabic Speech Dataset for Seasme (from MCV17)
Dataset Description
This dataset is a curated and preprocessed version of the Arabic (ar) subset from Mozilla Common Voice (MCV) 17.0. It has been specifically prepared for fine-tuning conversational speech models, with a primary focus on the Seasme-CSM model architecture. The dataset consists of audio clips in WAV format (24kHz, mono) and their corresponding transcripts, along with integer speaker IDs.
The original… See the full description on the dataset page: https://huggingface.co/datasets/MAdel121/Common-Voice-17-Arabic-for-Seasme-CSM-Finetuning.OpenWhistle-Classification-Finetuning
OpenWhistle Classification Finetuning Dataset
OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning is the public
classification finetuning dataset used for dolphin whistle identity
classification. It contains short whistle clips, whistle-level metadata,
fundamental-frequency tracks, rendered F0 spectrograms, and integer class
labels.
The main reviewer-facing subset is the balanced balanced config. It contains
six classes:
NSW_1 (label=0)
SW_Luna (label=1)
SW_Nana (label=2)… See the full description on the dataset page: https://huggingface.co/datasets/OpenWhistleNeurIPS26/OpenWhistle-Classification-Finetuning.fineTuningDataIndian_Englsih_SSML_dataset_for_orpheus_fine_tuningFYP_Fine_Tuning
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/shane062/FYP_Fine_Tuning.whisper-finetuningtts-finetuningMDbA_FineTuningTTS_finetuning_datasetarticulationGAN_finetuning_data
Dataset Card for "articulationGAN_finetuning_data"
More Information needed
whisper-finetuning-for-aseeMDbA-FineTuningV2nova_finetuningai-residency-whisper-fine-tuning-dataWhisper_FineTuning_KoWhisper_FineTuning_Suwhisper_finetuningfine-tuning-Jewish
Dataset Card for "fine-tuning-Jewish"
More Information needed
Zeta-Voice-ID-Vits-Finetuningtecholution_noun_finetuning_v1whisper-finetuning-testwhisper_finetuningwhisper-finetuning-audio-filesSpeech_to_fine_tuning
FeruzaSpeech_to_fine_tuning
A speech corpus of ⏱️ ~59.1 total hours of Uzbek audio paired with Latin‑script transcripts, intended for fine‑tuning ASR / speech‑to‑text models.
Dataset Details
Dataset Description
This dataset contains recordings of native Uzbek speakers reading a mix of classical literature excerpts and school‑level writing prompts:
001: Choliqushi (a novel by Rashod Nuri Guntekin, trans. by Mirzakalon Ismoiliy; first pub. Sept 1900).
002:… See the full description on the dataset page: https://huggingface.co/datasets/hostbot77/Speech_to_fine_tuning.stt_lowrank_finetuning
