datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wakeforge-hey-android-piper-tts
hey_android — Wake Word Synthetic Speech Dataset
Synthetic, augmented audio for training a small wake-word / keyword-spotting
model. Generated with piper_tts and local audio augmentation.
Classes
Label
Samples
background_noise
200
hey_android
378
unknown
1071
hey_android — the target wake phrase and close variants.
unknown — near-miss and unrelated short phrases.
background_noise — synthetic background noise.
Audio Specification… See the full description on the dataset page: https://huggingface.co/datasets/eoinedge/wakeforge-hey-android-piper-tts.Jarvis_Piper_TTS_Voicepiper-plus-tts-modelsnisan_kumru_piper_ttspiper_tts_kipermazlum_kiper_piper_ttspiper_tts_murat_eken_v0tricky-tts-piper-en-gbpiper-tts-dataset-wallpad2bvt2203_torkhov_30_pipertts_luxembourgishpiper-tts-dataset-wallpadpiper-tts-glasses-datasetpiper-tts-datasetpiper-tts-dataset-mystylepiper-hd-tts-datasetpiper-hd-tts-dataset2
