CoolFace
Datasetpublicgated

Sebssihakim/Clean_One_Speaker_SADA

Clean_One_Speaker_SADA A cleaned, single-speaker subset of the SADA (Saudi Audio Dataset for Arabic) corpus, derived from MahmoudIbrahim/100Hours-SADA22. Processing Follows the cleaning procedure from the Kaggle notebook Segmented Audio Data for Arabic Dialects (SADA): Removed rows with Unknown speaker age or gender. Removed rows whose dialect is More than 1 speaker, Unknown, or Notapplicable (every remaining segment has exactly one identified speaker).… See the full description on the dataset page: https://huggingface.co/datasets/Sebssihakim/Clean_One_Speaker_SADA.

sourceHugging Facecc-by-nc-sa-4.0updated 3mo agoView on Hugging Face
0likes11downloads
settings

This repository belongs to Sebssihakim on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameClean_One_Speaker_SADA
visibilitypublic
licencecc-by-nc-sa-4.0
gatedyes
ownerSebssihakim
Account settings
Sebssihakim/Clean_One_Speaker_SADA · CoolFace