CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01alakxender /dhivehi-audios-82-spk Dhivehi Synthetic Voice and Speech Augmentation Dataset This dataset is a multi-speaker dataset containing 1.26 million synthetic audio samples (~2,627 hours total). Each sample pairs a Dhivehi sentence with an augmented waveform, created through controlled synthesis, voice-cloning, and heavy acoustic perturbations. The dataset was generated to enable ASR, TTS, and voice-representation research in low-resource Dhivehi, focusing on robustness across pronunciation, prosody, and timbre… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-audios-82-spk.audioautomatic-speech-recognition1M<n<10M2 likes2.2k downloads11mo agoHugging Face02Serialtechlab /dhivehi-tts-preprocessedaudio10K<n<100K0 likes152 downloads7mo agoHugging Face03alakxender /dhivehi-conversations-turn Dhivehi Conversations (Turn-Based) This is an experimental synthetic dataset of turn-based Dhivehi conversations created for testing and fine-tuning dialogue models, text-to-speech (TTS), and multi-turn speaker-aware systems. This dataset is artificially constructed and not based on real conversations. It is intended for research experimentation only and may not always produce contextually accurate results. Dataset Source Derived from alakxender/voice-synthetic… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-conversations-turn.audiotext-to-speech10K<n<100K0 likes140 downloads1y agoHugging Face04Serialtechlab /dhivehi-tts-female-refined-splitaudio10K<n<100K0 likes81 downloads7mo agoHugging Face05alakxender /dhivehi-audio-kn Dhivehi Audio Dataset This is a Dhivehi (Maldivian) speech synthesis dataset with audio recordings, text transcriptions, and phonetic annotations. All recordings are by one speaker. Dataset Overview This dataset provides 4,170 high-quality audio samples in Dhivehi. Key Statistics Metric Value Total Audio Files 4,170 Total Duration 5.87 hours (352.0 minutes) Average Clip Length 5.06 seconds Total Words 30,068 Unique Phonemes 14190 Unique… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-audio-kn.audiotext-to-speech1K<n<10K0 likes59 downloads5mo agoHugging Face06alakxender /dhivehi-audio-casts Dhivehi-Audio-Casts Audio dataset with transcripts and speaker characteristics extracted from Qdrant database. Dataset Summary This dataset contains 3,300 audio recordings with corresponding transcripts and speaker characteristics including: Audio: WAV format audio files Sentence: Transcribed text in Dhivehi Speaker Demographics: Age, gender probabilities (male/female/child) Audio Characteristics: Arousal, dominance, valence scores Audio Length: Duration in seconds… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-audio-casts.audioautomatic-speech-recognition1K<n<10K0 likes38 downloads1y agoHugging Face07shiimi /dhivehi-audio-casts-processedaudio1K<n<10K0 likes28 downloads1y agoHugging Face08javaabu /dhivehi-shaafiu-speechgatedDhivehi Shaafiu Speech is a single speaker Dhivehi speech dataset created by [Javaabu Pvt. Ltd.](https://javaabu.com). The dataset contains around 16.5 hrs of text read by professional Maldivian narrator Muhammadh Shaafiu. The text used for the recordings were text scrapped from various Maldivian news websites.audioautomatic-speech-recognition1K<n<10K2 likes22 downloads3y agoHugging Face09Serialtechlab /dhivehi-javaabu-speech-parquetgatedaudio1K<n<10K0 likes22 downloads9mo agoHugging Face10Serialtechlab /dhivehi-tts-female-01gatedaudio10K<n<100K0 likes21 downloads10mo agoHugging Face11saajidha /dhivehi_speech_dataset Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Repository: (https://huggingface.co/datasets/saajidha/dhivehi_speech_dataset) Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Direct Use Out-of-Scope Use Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/saajidha/dhivehi_speech_dataset.audioautomatic-speech-recognitionn<1K0 likes18 downloads1y agoHugging Face12Devion333 /dhivehi-audio-1000-splitaudio1K<n<10K0 likes18 downloads1y agoHugging Face13javaabu /dhivehi-majlis-speechgatedDhivehi Majlis Speech is a Dhivehi speech dataset created from data annotated by [Javaabu Pvt. Ltd.](https://javaabu.com). The dataset contains around 10.5 hrs of speech collected from parliament sessions at The Peoples Majlis of Maldives (Maldivian Parliament) consisting of audio from different MPs from 6 different sessions.audioautomatic-speech-recognition1K<n<10K1 likes13 downloads3y agoHugging Face14Serialtechlab /dhivehi-tts-male-02-splitaudio1K<n<10K0 likes10 downloads10mo agoHugging Face15javaabu /dhivehi-khadheeja-speechgatedDhivehi Khadheeja Speech is a single speaker Dhivehi speech dataset created by [Javaabu Pvt. Ltd.](https://javaabu.com). The dataset contains around 20 hrs of text read by professional Maldivian narrator Khadheeja Faaz. The text used for the recordings were text scrapped from various Maldivian news websites.audioautomatic-speech-recognition1K<n<10K0 likes9 downloads2y agoHugging Face16chumputy /dhivehi-shaafiu-speech-trainaudio1K<n<10K0 likes7 downloads2y agoHugging Face17Serialtechlab /common-voice-dhivehi-malegatedaudio1K<n<10K0 likes7 downloads10mo agoHugging Face18alakxender /dhivehi-audios-ds2gated Dhivehi Audio Dataset 2 A quality-filtered Dhivehi speech dataset with three subsets (bronze, silver, gold), each representing a progressively stricter quality threshold. Subsets Subset MOS CTC gc WER Train Test bronze ≥3.0 ≥0.70 — 65,838 7,316 silver ≥3.0 ≥0.70 ≤0.30 43,010 4,779 gold ≥3.5 ≥0.75 ≤0.20 7,058 785 MOS — Mean Opinion Score (1–4), subjective listening quality rating CTC gc — CTC forced-alignment geo-confidence (0–1), measures… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-audios-ds2.audioautomatic-speech-recognition100K<n<1M0 likes7 downloads4mo agoHugging Face19alakxender /dhivehi-audios-ds1gated Dhivehi Audio Dataset 1 A quality-filtered Dhivehi speech dataset with three subsets (bronze, silver, gold), each representing a progressively stricter quality threshold. Subsets Subset MOS CTC gc WER Train Test bronze ≥3.0 ≥0.70 — 224,321 24,925 silver ≥3.0 ≥0.70 ≤0.30 131,389 14,599 gold ≥3.5 ≥0.75 ≤0.20 15,535 1,727 MOS — Mean Opinion Score (1–4), subjective listening quality rating CTC gc — CTC forced-alignment geo-confidence (0–1)… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-audios-ds1.audioautomatic-speech-recognition100K<n<1M0 likes6 downloads4mo agoHugging Face20Serialtechlab /dhivehi-tts-male-03-splitaudio10K<n<100K0 likes5 downloads10mo agoHugging Face21Serialtechlab /dhivehi-tts-female-01-split Dataset Card for "dhivehi-tts-female-01-split" More Information needed audio1K<n<10K0 likes3 downloads10mo agoHugging Face22inshal /Quran-dhivehi-translation-audio-verse-by-versegatedaudio1K<n<10K0 likes3 downloads6mo agoHugging Face23Serialtechlab /dhivehi-tts-male-refined-splitgatedaudio10K<n<100K0 likes2 downloads10mo agoHugging Face24Serialtechlab /dhivehi-mms-v5-combinedgatedaudio1K<n<10K2 likes2 downloads9mo agoHugging Face25Serialtechlab /dhivehi-tts-combined-splitgatedaudio10K<n<100K0 likes2 downloads7mo agoHugging Face26javaabu /scaling_dhivehi_sttgatedaudio100K<n<1M0 likes2 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.