datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
TOSD
Dataset Card for Tamazight Open Speech Dataset
This dataset provides a parsed, formatted, and ready-to-use Amazigh Voice Dataset. It contains voice recordings and corresponding text transcripts in Standard Moroccan Amazigh (ⵜⴰⵎⴰⵣⵉⵖⵜ ⵜⴰⵏⴰⵡⴰⵢⵜ ⵜⴰⵎⵓⵔⴰⴽⵓⵛⵜ) intended for training Automatic Speech Recognition (ASR) and Text-to-Speech (TTS) models.
This specific repository is published by a collaborator. You may visit the raw dataset repository which has additional dataset that hasn't… See the full description on the dataset page: https://huggingface.co/datasets/Tamazight-NLP/TOSD.Tosh314
