datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
yali-lower
Yali-Lower: Isolated Mandarin Syllables with Full Tone Coverage (lower-pitch version)
Yali语音:带完整声调的孤立汉语音节录音库
These syllables were recorded in context by
Cheng Ya Li for the Gradint program
in June 2008. Sound editing was done by Cameron Wong
(developer of the Ekho
speech synthesizer) and Silas S. Brown, using
Audacity and Praat.
这些音节由程雅丽(Cheng Ya Li,汉字待确认)于2008年6月为
Gradint程序录制。音频编辑由黄冠能(Cameron
Wong,Ekho语音合成器的开
发者)和Silas S. Brown(赛乐思)使用Audacity和Praat完成。… See the full description on the dataset page: https://huggingface.co/datasets/real-ssb22/yali-lower.Adaption-low-resource-audio
Adaption Low-Resource Audio
A low-resource-language subset of
Reubencf/PolyglotAudio,
remastered with Adaption's Adaptive Data
platform. Each row carries the original Tatoeba-derived audio clip
alongside sharpened enhanced_prompt / enhanced_completion columns
so the data is ready for speech-model fine-tuning and evaluation on
languages that are typically under-represented in open ASR/TTS corpora.
Dataset size
3,704 rows of paired audio + text, spanning 10 languages… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/Adaption-low-resource-audio.low-decoder
low-decoding
Author: Surpem
This dataset contains 1200 unique, clean synthetic audio signals representing decoded text commands.
The audio signals represent synthesized Morse Code message blocks.
Dataset Structure
id: A unique UUID string.
audio: The audio wav bytes (16kHz Mono).
text: The decoded string transcription.
Dataset Level: LOW
Low: Slow WPM (~12 WPM), high signal-to-noise ratio (clean), short command strings.
Medium: Fast WPM (~24 WPM), background… See the full description on the dataset page: https://huggingface.co/datasets/Surpem/low-decoder.
