CoolFace
Datasetpublic

PerSets/filimo-persian-asr

This dataset consists of about 400 hours of audio extracted from various Filimo videos in the Persian language. Note: This dataset contains raw, unvalidated transcriptions. Users are advised to: 1. Perform their own quality assessment 2. Create their own train/validation/test splits based on their specific needs 3. Validate a subset of the data if needed for their use case

sourceHugging Facecc0-1.0updated 2y agoView on Hugging Face
7likes386downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
PerSets/filimo-persian-asr · CoolFace