snapwre/amharic-speech
Dataset.ET Amharic Speech — v0.2.0 51.547 hours · 16,866 clips · 493 speakers · 15,443 distinct prompts Dataset Summary Read speech in Amharic, crowdsourced from volunteer contributors in Ethiopia through a Telegram bot, peer-validated by other contributors, and screened acoustically before release. Amharic has very little open speech data; this corpus exists to change that. Contributors read a displayed prompt aloud, other contributors listen and vote on whether… See the full description on the dataset page: https://huggingface.co/datasets/snapwre/amharic-speech.
Remove shards superseded by this release
Add files using upload-large-folder tool
Dataset.ET Amharic Speech v0.1.0 — 22.7h, 7,405 clips, 320 speakers
Dataset.ET Amharic Speech v0.1.0 — 22.7h, 7,405 clips, 320 speakers
Dataset.ET Amharic Speech v0.1.0 — 22.7h, 7,405 clips, 320 speakers
initial commit
