datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
palavras_pt_faber_10K_cadu_100Cadupalavras_pt_faber_15K_cadu_150BangChan_trainYunjin_trainpalavras_pt_cadu_150cadencecadence_speachsoch-tts
soch-2h TTS Dataset
Generated by dataset-maker.
Samples: 984
Format: HuggingFace audiofolder (Qwen 3 TTS compatible)
Columns:
audio — audio segment (decoded Audio feature, with player in viewer)
text — transcript
ref_audio — string path to reference speaker audio (data/ref_audio.wav)
palavras_pt_cadu_100
