PerSets/filimo-persian-asr
This dataset consists of about 400 hours of audio extracted from various Filimo videos in the Persian language. Note: This dataset contains raw, unvalidated transcriptions. Users are advised to: 1. Perform their own quality assessment 2. Create their own train/validation/test splits based on their specific needs 3. Validate a subset of the data if needed for their use case
7382
