alakxender/dhivehi-audios-ds2
Dhivehi Audio Dataset 2 A quality-filtered Dhivehi speech dataset with three subsets (bronze, silver, gold), each representing a progressively stricter quality threshold. Subsets Subset MOS CTC gc WER Train Test bronze ≥3.0 ≥0.70 — 65,838 7,316 silver ≥3.0 ≥0.70 ≤0.30 43,010 4,779 gold ≥3.5 ≥0.75 ≤0.20 7,058 785 MOS — Mean Opinion Score (1–4), subjective listening quality rating CTC gc — CTC forced-alignment geo-confidence (0–1), measures… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-audios-ds2.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face