Alirezav99/ganjoor
Ganjoor Persian Speech Dataset Dataset Description This dataset contains Persian speech recordings from Ganjoor.ir, segmented based on Persian poetry verses. The audio files have been processed and segmented into manageable chunks suitable for speech-to-text (STT) training and evaluation. Dataset Summary Language: Persian (Farsi) Domain: Persian poetry (classical and contemporary) Task: Automatic Speech Recognition (ASR) Format: MP3 audio files… See the full description on the dataset page: https://huggingface.co/datasets/Alirezav99/ganjoor.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face