datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mishkat-quran-audio
تلاوةُ الحصريّ — مرآةٌ لعمل مشكاة بلا إنترنت
Al-Ḥuṣarī recitation — an offline mirror for Mishkat
العربيّة أوّلاً، ثمّ الإنجليزيّة. · Arabic first, then English.
ما هذا؟
ملفّاتُ تلاوةٍ آيةً آيةً للشيخ محمود خليل الحصريّ، منسوخةٌ كما هي من
everyayah.com بلا قصٍّ ولا إعادةِ ترميز، ليعمل بها
تطبيق مشكاة — الاستماعُ والتلقينُ — بلا اتّصالٍ
بالإنترنت.
ولا نصَّ قرآنٍ في هذا المستودع ولا تفسير — صوتٌ فقط، ومانيفستٌ يصفه.
القارئان — ولكلٍّ… See the full description on the dataset page: https://huggingface.co/datasets/emadjumaah/mishkat-quran-audio.SPIRE_EMA_CORPUSThis corpus contains paired data of speech, articulatory movements and phonemes. There are 38 speakers in the corpus, each with 460 utterances.
The raw audio files are in audios.zip. The ema data and preprocessed data is stored in processed.zip. The processed data can be loaded with pytorch and has the following keys -
ema_raw : The raw ema data
ema_clipped : The ema data after trimming using being-end time stamps
ema_trimmed_and_normalised_with_6_articulators: The ema data after trimming… See the full description on the dataset page: https://huggingface.co/datasets/SpireLab/SPIRE_EMA_CORPUS.japaSPIRE_EMA_CORPUSThis corpus contains paired data of speech, articulatory movements and phonemes. There are 38 speakers in the corpus, each with 460 utterances.
The raw audio files are in audios.zip. The ema data and preprocessed data is stored in processed.zip. The processed data can be loaded with pytorch and has the following keys -
ema_raw : The raw ema data
ema_clipped : The ema data after trimming using being-end time stamps
ema_trimmed_and_normalised_with_6_articulators: The ema data after trimming… See the full description on the dataset page: https://huggingface.co/datasets/viks66/SPIRE_EMA_CORPUS.EMAGEflemish-emailEmarati-Speech-Datasete_mask_activ2e_mask_edu
