CoolFace
Datasetpublic

suleiman2003/W_hausa_v3

Cleaned Hausa Speech Dataset v3 A cleaned and processed Hausa speech dataset built from multiple open-source Hugging Face datasets. Dataset Description This dataset contains cleaned, normalized, and deduplicated Hausa speech audio with aligned transcriptions. All audio is: Sample rate: 16,000 Hz (mono) Format: FLAC (lossless, embedded in Parquet) Duration range: 1–30 seconds per clip Loudness normalized: -20 dBFS RMS VAD trimmed: Non-speech segments removed with… See the full description on the dataset page: https://huggingface.co/datasets/suleiman2003/W_hausa_v3.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes329downloads
settings

This repository belongs to suleiman2003 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameW_hausa_v3
visibilitypublic
licencecc-by-4.0
gatedno
ownersuleiman2003
Account settings
suleiman2003/W_hausa_v3 · CoolFace