CoolFace
Datasetpublic

oddadmix/msa-omnivoice-tts-v1

MSA-OmniVoice-v1 Dataset Description MSA-OmniVoice-v1 is a 50-hour synthetic Modern Standard Arabic (MSA) speech dataset generated using OmniVoice. The dataset contains high-quality synthetic speech from a single speaker paired with fully diacritized (تشكيل) transcripts. It is intended for training and fine-tuning Arabic speech models, including Text-to-Speech (TTS), Automatic Speech Recognition (ASR), speech representation learning, and alignment tasks.… See the full description on the dataset page: https://huggingface.co/datasets/oddadmix/msa-omnivoice-tts-v1.

sourceHugging Faceupdated 3mo agoView on Hugging Face
1likes341downloads
3 commits on main
bdbfed93mo ago

Update README.md

oddadmix
8f76da73mo ago

Upload 30853 MSA TTS samples

oddadmix
63428c83mo ago

initial commit

oddadmix