CoolFace
Datasetpublic

BSC-LT/distilled-yodas-spanish

Distilled YODAS Spanish is a high-quality subset of the Spanish portion of the YouTube-Oriented Dataset for Audio and Speech (YODAS). While the full YODAS corpus contains over 37,000 hours of Spanish speech across 43 million files, this dataset provides a distilled version of approximately 8,000 validated hours.

sourceHugging Facecc-by-3.0updated 10mo agoView on Hugging Face
5likes124downloads

No commit history came back for main. The revision may not exist, or the source declined the request.