Thomcles/YodaLingua-Farsi
YodaLingua-Farsi YodaLingua is a high-quality speech dataset designed for training text-to-speech (TTS) systems, ASR models, and any application requiring clean, well-aligned audio–text pairs.This release contains the Farsi portion of the multilingual YodaLingua collection. 🧾 Dataset Overview Property Value Total clips 23,419 audio–transcription pairs Total duration 72 hours Speakers 678 distinct speakers Audio format MP3 • mono • 24 kHz • 16-bit… See the full description on the dataset page: https://huggingface.co/datasets/Thomcles/YodaLingua-Farsi.
This repository belongs to Thomcles on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
