CoolFace
Datasetpublic

BSC-LT/distilled-yodas-spanish

Distilled YODAS Spanish is a high-quality subset of the Spanish portion of the YouTube-Oriented Dataset for Audio and Speech (YODAS). While the full YODAS corpus contains over 37,000 hours of Spanish speech across 43 million files, this dataset provides a distilled version of approximately 8,000 validated hours.

sourceHugging Facecc-by-3.0updated 10mo agoView on Hugging Face
5likes124downloads
settings

This repository belongs to BSC-LT on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namedistilled-yodas-spanish
visibilitypublic
licencecc-by-3.0
gatedno
ownerBSC-LT
Account settings
BSC-LT/distilled-yodas-spanish · CoolFace