speech-audio
Multitask-National-Speech-Corpus-v1-extendTumbuka_Text-Speech_Audiochinese_speech_sft_cosy_audioThis dataset use Cosy-voice 2 0.5B to generate speech audio from https://huggingface.co/datasets/TigerResearch/sft_zh.
https://huggingface.co/datasets/JerryAGENDD/voiceprint_librispeech_other_test is used as audio input prompt.
Deepfake-Eval-2024-audiogerman-golden-audio_speech-IPA
🌟 German Golden Speech & IPA Corpus (FLEURS + Multilingual TEDx)
An ultra-clean, high-standard curated German speech dataset combining Google FLEURS (de_de) and Multilingual TEDx German (mTEDx), fully embedded with 16kHz WAV audio bytes, normalized orthographic text, and pre-computed International Phonetic Alphabet (IPA) transcriptions.
📊 Dataset Summary
Total Samples: 1,354 high-quality audio recordings.
Total Size: ~419 MB (Compressed Parquet format).
Audio… See the full description on the dataset page: https://huggingface.co/datasets/q1805/german-golden-audio_speech-IPA.parlertts-pony-speech-audio
