CoolFace
Datasetpublic

oddadmix/msa-omnivoice-tts-v1

MSA-OmniVoice-v1 Dataset Description MSA-OmniVoice-v1 is a 50-hour synthetic Modern Standard Arabic (MSA) speech dataset generated using OmniVoice. The dataset contains high-quality synthetic speech from a single speaker paired with fully diacritized (تشكيل) transcripts. It is intended for training and fine-tuning Arabic speech models, including Text-to-Speech (TTS), Automatic Speech Recognition (ASR), speech representation learning, and alignment tasks.… See the full description on the dataset page: https://huggingface.co/datasets/oddadmix/msa-omnivoice-tts-v1.

sourceHugging Faceupdated 3mo agoView on Hugging Face
1likes341downloads

oddadmix/msa-omnivoice-tts-v1 · main · files are served by the source, never re-hosted here