pymmdrza/Common-Voice-Speech-26.0-Persian-Clean
Persian Common Voice Clean Dataset This dataset is a cleaned and prepared subset of the Persian (فارسی - fa) portion of Mozilla Common Voice Scripted Speech, based on cv-corpus-26.0-2026-06-12. The cleaned release contains 34,134 audio clips, representing approximately 43.105 hours of speech, equal to 2,586.303 minutes. The clips are associated with approximately 34,134 validated Persian sentences and come from 3,791 speakers. The original Persian Common Voice release contains… See the full description on the dataset page: https://huggingface.co/datasets/pymmdrza/Common-Voice-Speech-26.0-Persian-Clean.
Update README.md
Update README.md
Update README.md
Publish standard audio Parquet dataset
Update dataset_info.json
Update README.md
Add dataset info
Update dataset card
Update README.md
Create README.md
Upload folder using huggingface_hub (part 16)
Upload folder using huggingface_hub (part 15)
Upload folder using huggingface_hub (part 14)
Upload folder using huggingface_hub (part 13)
Upload folder using huggingface_hub (part 12)
Upload folder using huggingface_hub (part 11)
Upload folder using huggingface_hub (part 10)
Upload folder using huggingface_hub (part 9)
Upload folder using huggingface_hub (part 8)
Upload folder using huggingface_hub (part 7)
Upload folder using huggingface_hub (part 6)
Upload folder using huggingface_hub (part 5)
Upload folder using huggingface_hub (part 4)
Upload folder using huggingface_hub (part 3)
Upload folder using huggingface_hub (part 2)
Upload folder using huggingface_hub
initial commit
