CoolFace
Datasetpublic

snapwre/amharic-speech

Dataset.ET Amharic Speech — v0.2.0 51.547 hours · 16,866 clips · 493 speakers · 15,443 distinct prompts Dataset Summary Read speech in Amharic, crowdsourced from volunteer contributors in Ethiopia through a Telegram bot, peer-validated by other contributors, and screened acoustically before release. Amharic has very little open speech data; this corpus exists to change that. Contributors read a displayed prompt aloud, other contributors listen and vote on whether… See the full description on the dataset page: https://huggingface.co/datasets/snapwre/amharic-speech.

sourceHugging Facecc-by-4.0updated 27d agoView on Hugging Face
25likes1.2kdownloads
6 commits on main
4861d6427d ago

Remove shards superseded by this release

Chapimenge
8c6d52927d ago

Add files using upload-large-folder tool

Chapimenge
9ca01831mo ago

Dataset.ET Amharic Speech v0.1.0 — 22.7h, 7,405 clips, 320 speakers

Chapimenge
91848b71mo ago

Dataset.ET Amharic Speech v0.1.0 — 22.7h, 7,405 clips, 320 speakers

Chapimenge
d35c0a61mo ago

Dataset.ET Amharic Speech v0.1.0 — 22.7h, 7,405 clips, 320 speakers

Chapimenge
b7c735f1mo ago

initial commit

Chapimenge