datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bible_tts_hausa
Dataset Card for BibleTTS Hausa
Dataset Summary
BibleTTS is a large high-quality open Text-to-Speech dataset with up to 80 hours of single speaker, studio quality 48kHz recordings.
This is a Hausa part of the dataset. Aligned hours: 86.6, aligned verses: 40,603.
Languages
Hausa
Dataset Structure
Data Fields
audio: audio path
sentence: transcription of the audio
locale: always set to ha
book: 3-char book encoding
verse: verse id… See the full description on the dataset page: https://huggingface.co/datasets/vpetukhov/bible_tts_hausa.W_hausa_v3
Cleaned Hausa Speech Dataset v3
A cleaned and processed Hausa speech dataset built from multiple open-source Hugging Face datasets.
Dataset Description
This dataset contains cleaned, normalized, and deduplicated Hausa speech audio with aligned transcriptions. All audio is:
Sample rate: 16,000 Hz (mono)
Format: FLAC (lossless, embedded in Parquet)
Duration range: 1–30 seconds per clip
Loudness normalized: -20 dBFS RMS
VAD trimmed: Non-speech segments removed with… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/W_hausa_v3.W_hausa_v7
Unified Hausa Speech Dataset v5
Dataset Description
A large-scale, cleaned, deduplicated, and quality-filtered Hausa speech dataset compiled from multiple open-source collections. Designed for Text-to-Speech (TTS) and Automatic Speech Recognition (ASR) research.
All audio is 16 kHz mono FLAC, silence-trimmed, loudness-normalized to -20 dBFS, and sorted by speaker_id so that all clips from the same speaker appear consecutively.
Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/W_hausa_v7.Hausa-Synthetic-ASR-Dataset-XTTSSynthetic Hausa ASR dataset generated using a fine-tuned version of the XTTS-v2 model.
Sample rate: 24kHz.
Total duration: 574 hours.
fongbe-hausa-asr-dataset
Fongbe-Hausa ASR Dataset (Semi-Supervised)
This dataset provides ~6,770 audio-transcription pairs for Fongbe (fon) and Hausa (hau). It was created using a semi-supervised pipeline to convert long-form video content into a training-ready format for Automatic Speech Recognition (ASR).
Dataset Details
Total Examples: 6,770
Audio Format: WAV (16kHz, Mono)
Languages: Fongbe (Benin), Hausa (Nigeria/West Africa)
Annotation: Semi-supervised (Machine-generated labels)
License:… See the full description on the dataset page: https://huggingface.co/datasets/Professor/fongbe-hausa-asr-dataset.unified-hausa-speech
Unified Hausa Speech Dataset v5
Dataset Description
A large-scale, cleaned, deduplicated, and quality-filtered Hausa speech dataset compiled from 6 open-source collections. Designed for Text-to-Speech (TTS) and Automatic Speech Recognition (ASR) research on one of Africa's most widely spoken languages.
Hausa (ISO 639-1: ha) is a Chadic language spoken by over 80 million people across West and Central Africa — primarily in Nigeria and Niger, and as a trade language… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/unified-hausa-speech.hausa-speech-transcribed
Hausa Spontaneous Speech, Transcribed — Silencio
Spontaneous Hausa with human-validated transcription and word-level alignment. 49 clips from 23 distinct speakers in Nigeria, most of them from Kano and Katsina, the core of the Kano-standard Hausa area. Transcripts are in standard Boko (Latin) orthography, with the hooked consonants ɓ, ɗ, ƙ and ƴ.
Hours
0.36
Clips
49
Speakers
23
Countries
1
Speaker origin regions
5
Native Hausa speakers (declared)
20 of 23… See the full description on the dataset page: https://huggingface.co/datasets/SilencioNetwork/hausa-speech-transcribed.hausa_voice_dataset
Dataset Card for "hausa_voice_dataset"
Dataset Overview
Dataset Name: Hausa Voice Dataset
Description: This dataset contains Hausa language audio samples from Common Voice. The dataset includes audio files and their corresponding transcriptions, designed for text-to-speech (TTS) and automatic speech recognition (ASR) research and applications.
Dataset Structure
Configs:
default
Data Files:
Split: train
Dataset Info:
Features:
audio: Audio file (mono… See the full description on the dataset page: https://huggingface.co/datasets/mide7x/hausa_voice_dataset.hausa_long_voice_dataset
Dataset Card for "hausa_long_voice_dataset"
Dataset Overview
Dataset Name: Hausa Long Voice Dataset
Description: This dataset contains merged Hausa language audio samples from Common Voice. Audio files from the same speaker have been concatenated to create longer audio samples with their corresponding transcriptions, designed for text-to-speech (TTS) training where longer sequences are beneficial.
Dataset Structure
Configs:
default
Data Files:
Split: train… See the full description on the dataset page: https://huggingface.co/datasets/mide7x/hausa_long_voice_dataset.9javoice-hausa
9jaVoice Consent-1
74 clips. 19.1 minutes, which is 0.3 hours. 6 speakers. Hausa. FLAC, 48 kHz, mono, 16-bit. CC BY-NC 4.0.
Read-aloud speech, recorded by paid contributors on their own phones. Use it for evaluation, for fine-tuning, and as a reference set when you want to find out whether a model handles Nigerian speech at all. It is too small to pretrain on and we are not going to pretend otherwise.
Every clip here carries its own consent record. The contributor ticked an… See the full description on the dataset page: https://huggingface.co/datasets/9jatesters/9javoice-hausa.Hausa-Speech-Dataset
🎧 Hausa Speech Dataset
The Hausa Speech Dataset is a structured and high-quality speech audio dataset developed to support modern AI systems requiring diverse audio data and reliable voice data. It includes 174 hours of recordings distributed across 733 files, available in MP3 and WAV formats, with a total size of 362 MB. This carefully curated audio dataset ensures balanced speaker representation, with 52% female and 48% male speakers, and a broad age distribution from 18 to 50+… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Hausa-Speech-Dataset.Hausa-Synthetic-ASR-Dataset-YourTTSSynthetic Hausa ASR dataset generated using a fine-tuned version of the YourTTS model.
Sample rate: 24kHz.
Total duration: 993 hours.
