CoolFace
Datasetpublic

C1Tech/Persian-ASR-Benchmark

This dataset consists of 3 hours of 16kHz audio collected from diverse environments to better represent real-world scenarios. The recordings were sourced from audiobooks, YouTube, and other public sources, ensuring a wide variety of speech styles and acoustic conditions. One key advantage of this dataset is that it was collected from recent sources within the last few months, ensuring no overlap with training data and fairness for evaluating other STT models. To enable a robust and fair… See the full description on the dataset page: https://huggingface.co/datasets/C1Tech/Persian-ASR-Benchmark.

sourceHugging Faceupdated 2mo agoView on Hugging Face
4likes244downloads

No commit history came back for main. The revision may not exist, or the source declined the request.