datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
quranic-universal-ayahs
Qur'anic Universal Ayahs
Qur'anic Universal Audio (QUA) is a project that unifies recitations on the internet and generates timing data using forced alignment — community-verified results and constantly expanding dataset.
This dataset pairs ayah by ayah audio with word-level timestamps, DigitalKhatt letter-animation timestamps, and waqf-aware segment data. Repeated words are preserved in text_uthmani and word_timestamps, so the row reflects what the reciter… See the full description on the dataset page: https://huggingface.co/datasets/QUD-Technologies/quranic-universal-ayahs.quranic-universal-ayahs
Qur'anic Universal Ayahs
Qur'anic Universal Audio (QUA) is a project that unifies recitations on the internet and generates timing data using forced alignment — community-verified results and constantly expanding dataset.
This dataset pairs ayah by ayah audio with word-level timestamps, DigitalKhatt letter-animation timestamps, and waqf-aware segment data. Repeated words are preserved in text_uthmani and word_timestamps, so the row reflects what the reciter… See the full description on the dataset page: https://huggingface.co/datasets/hetchyy/quranic-universal-ayahs.eval-universal-3-pro-eka-hard-20260408-1924
Evaluation Results: universal-3-pro
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
assemblyai/universal-3-pro
43.36%
33.70%
Source Data
Evaluation Dataset: Trelis/eka-hard
Model Evaluated: assemblyai/universal-3-pro
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate for this… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-universal-3-pro-eka-hard-20260408-1924.eval-universal-3-pro-medical-terms-2025-20260408-1928
Evaluation Results: universal-3-pro
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
assemblyai/universal-3-pro
6.91%
2.40%
Source Data
Evaluation Dataset: Trelis/medical-terms-2025
Model Evaluated: assemblyai/universal-3-pro
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate for… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-universal-3-pro-medical-terms-2025-20260408-1928.eval-universal-3-pro-multimed-hard-20260408-1933
Evaluation Results: universal-3-pro
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
assemblyai/universal-3-pro
12.53%
10.02%
Source Data
Evaluation Dataset: Trelis/multimed-hard
Model Evaluated: assemblyai/universal-3-pro
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate for… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-universal-3-pro-multimed-hard-20260408-1933.
