CoolFace
Datasetpublic

q1805/german-golden-audio_speech-IPA

🌟 German Golden Speech & IPA Corpus (FLEURS + Multilingual TEDx) An ultra-clean, high-standard curated German speech dataset combining Google FLEURS (de_de) and Multilingual TEDx German (mTEDx), fully embedded with 16kHz WAV audio bytes, normalized orthographic text, and pre-computed International Phonetic Alphabet (IPA) transcriptions. 📊 Dataset Summary Total Samples: 1,354 high-quality audio recordings. Total Size: ~419 MB (Compressed Parquet format). Audio… See the full description on the dataset page: https://huggingface.co/datasets/q1805/german-golden-audio_speech-IPA.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
0likes170downloads
5 commits on main
50a26fb1mo ago

Update README.md

q1805
fe21de71mo ago

Upload folder using huggingface_hub

q1805
2e294c11mo ago

Upload dataset

q1805
a4f3e061mo ago

Upload dataset

q1805
5a965011mo ago

initial commit

q1805