datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CosyVoice2-SparkTTSspark_ttsSparkTTS_Wang_Leehom_Ad_wavspark_tts_knn_vc_pathological
Dataset Overview
Total Samples: 785
Total Duration: 3006.46 seconds (50.11 minutes)
Speakers: 8 speakers
Corpora: TORGO, UA-Speech, LibriSpeech
Sample Rate: 16kHz (KNN-VC output rate, native Spark compatibility)
Audio Format: WAV
SparkTTS_Male_Ad_wavSparkTTS_Mavuika_Ad_wavSparkTTS_Female_Ad_wav
