CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01vpetukhov /bible_tts_hausa Dataset Card for BibleTTS Hausa Dataset Summary BibleTTS is a large high-quality open Text-to-Speech dataset with up to 80 hours of single speaker, studio quality 48kHz recordings. This is a Hausa part of the dataset. Aligned hours: 86.6, aligned verses: 40,603. Languages Hausa Dataset Structure Data Fields audio: audio path sentence: transcription of the audio locale: always set to ha book: 3-char book encoding verse: verse id… See the full description on the dataset page: https://huggingface.co/datasets/vpetukhov/bible_tts_hausa.textautomatic-speech-recognition10K<n<100K7 likes528 downloads4y agoHugging Face02suleiman2003 /W_hausa_v3 Cleaned Hausa Speech Dataset v3 A cleaned and processed Hausa speech dataset built from multiple open-source Hugging Face datasets. Dataset Description This dataset contains cleaned, normalized, and deduplicated Hausa speech audio with aligned transcriptions. All audio is: Sample rate: 16,000 Hz (mono) Format: FLAC (lossless, embedded in Parquet) Duration range: 1–30 seconds per clip Loudness normalized: -20 dBFS RMS VAD trimmed: Non-speech segments removed with… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/W_hausa_v3.audioautomatic-speech-recognition100K<n<1M0 likes327 downloads2mo agoHugging Face03suleiman2003 /W_hausa_v7 Unified Hausa Speech Dataset v5 Dataset Description A large-scale, cleaned, deduplicated, and quality-filtered Hausa speech dataset compiled from multiple open-source collections. Designed for Text-to-Speech (TTS) and Automatic Speech Recognition (ASR) research. All audio is 16 kHz mono FLAC, silence-trimmed, loudness-normalized to -20 dBFS, and sorted by speaker_id so that all clips from the same speaker appear consecutively. Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/W_hausa_v7.audiotext-to-speech100K<n<1M0 likes206 downloads9d agoHugging Face04CLEAR-Global /Hausa-Synthetic-ASR-Dataset-XTTSgatedSynthetic Hausa ASR dataset generated using a fine-tuned version of the XTTS-v2 model. Sample rate: 24kHz. Total duration: 574 hours. audioautomatic-speech-recognition100K<n<1M1 likes146 downloads1y agoHugging Face05Professor /fongbe-hausa-asr-dataset Fongbe-Hausa ASR Dataset (Semi-Supervised) This dataset provides ~6,770 audio-transcription pairs for Fongbe (fon) and Hausa (hau). It was created using a semi-supervised pipeline to convert long-form video content into a training-ready format for Automatic Speech Recognition (ASR). Dataset Details Total Examples: 6,770 Audio Format: WAV (16kHz, Mono) Languages: Fongbe (Benin), Hausa (Nigeria/West Africa) Annotation: Semi-supervised (Machine-generated labels) License:… See the full description on the dataset page: https://huggingface.co/datasets/Professor/fongbe-hausa-asr-dataset.audioautomatic-speech-recognition1K<n<10K0 likes96 downloads7mo agoHugging Face06suleiman2003 /unified-hausa-speech Unified Hausa Speech Dataset v5 Dataset Description A large-scale, cleaned, deduplicated, and quality-filtered Hausa speech dataset compiled from 6 open-source collections. Designed for Text-to-Speech (TTS) and Automatic Speech Recognition (ASR) research on one of Africa's most widely spoken languages. Hausa (ISO 639-1: ha) is a Chadic language spoken by over 80 million people across West and Central Africa — primarily in Nigeria and Niger, and as a trade language… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/unified-hausa-speech.audiotext-to-speech100K<n<1M0 likes72 downloads1mo agoHugging Face07SilencioNetwork /hausa-speech-transcribed Hausa Spontaneous Speech, Transcribed — Silencio Spontaneous Hausa with human-validated transcription and word-level alignment. 49 clips from 23 distinct speakers in Nigeria, most of them from Kano and Katsina, the core of the Kano-standard Hausa area. Transcripts are in standard Boko (Latin) orthography, with the hooked consonants ɓ, ɗ, ƙ and ƴ. Hours 0.36 Clips 49 Speakers 23 Countries 1 Speaker origin regions 5 Native Hausa speakers (declared) 20 of 23… See the full description on the dataset page: https://huggingface.co/datasets/SilencioNetwork/hausa-speech-transcribed.audioautomatic-speech-recognitionn<1K0 likes46 downloads2d agoHugging Face08mide7x /hausa_voice_dataset Dataset Card for "hausa_voice_dataset" Dataset Overview Dataset Name: Hausa Voice Dataset Description: This dataset contains Hausa language audio samples from Common Voice. The dataset includes audio files and their corresponding transcriptions, designed for text-to-speech (TTS) and automatic speech recognition (ASR) research and applications. Dataset Structure Configs: default Data Files: Split: train Dataset Info: Features: audio: Audio file (mono… See the full description on the dataset page: https://huggingface.co/datasets/mide7x/hausa_voice_dataset.audioautomatic-speech-recognition1K<n<10K1 likes41 downloads1y agoHugging Face09mide7x /hausa_long_voice_dataset Dataset Card for "hausa_long_voice_dataset" Dataset Overview Dataset Name: Hausa Long Voice Dataset Description: This dataset contains merged Hausa language audio samples from Common Voice. Audio files from the same speaker have been concatenated to create longer audio samples with their corresponding transcriptions, designed for text-to-speech (TTS) training where longer sequences are beneficial. Dataset Structure Configs: default Data Files: Split: train… See the full description on the dataset page: https://huggingface.co/datasets/mide7x/hausa_long_voice_dataset.audioautomatic-speech-recognitionn<1K0 likes33 downloads1y agoHugging Face109jatesters /9javoice-hausagated 9jaVoice Consent-1 74 clips. 19.1 minutes, which is 0.3 hours. 6 speakers. Hausa. FLAC, 48 kHz, mono, 16-bit. CC BY-NC 4.0. Read-aloud speech, recorded by paid contributors on their own phones. Use it for evaluation, for fine-tuning, and as a reference set when you want to find out whether a model handles Nigerian speech at all. It is too small to pretrain on and we are not going to pretend otherwise. Every clip here carries its own consent record. The contributor ticked an… See the full description on the dataset page: https://huggingface.co/datasets/9jatesters/9javoice-hausa.audioautomatic-speech-recognitionn<1K0 likes18 downloads2d agoHugging Face11Speech-data /Hausa-Speech-Dataset 🎧 Hausa Speech Dataset The Hausa Speech Dataset is a structured and high-quality speech audio dataset developed to support modern AI systems requiring diverse audio data and reliable voice data. It includes 174 hours of recordings distributed across 733 files, available in MP3 and WAV formats, with a total size of 362 MB. This carefully curated audio dataset ensures balanced speaker representation, with 52% female and 48% male speakers, and a broad age distribution from 18 to 50+… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Hausa-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes15 downloads6mo agoHugging Face12CLEAR-Global /Hausa-Synthetic-ASR-Dataset-YourTTSgatedSynthetic Hausa ASR dataset generated using a fine-tuned version of the YourTTS model. Sample rate: 24kHz. Total duration: 993 hours. audioautomatic-speech-recognition100K<n<1M0 likes12 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.