CoolFace
Datasetpublicgated

ErfanRou/persian-youtube-330h-soniox

persian-youtube-330h-soniox Persian conversational / voice-search / meeting ASR fine-tuning set built from YouTube audio, labelled by Soniox stt-async-v5. Machine-labelled. Not ground truth. Built 2026-09-22 by ytcrawl/build_dataset.py. What this is 74,538 training segments (331.5 h) and 7,442 validation segments (32.9 h), 16 kHz mono FLAC, 5–28 s each, cut from 2,708 videos on 106 channels across 9 domains. Validation is held out by channel (a creator is never on… See the full description on the dataset page: https://huggingface.co/datasets/ErfanRou/persian-youtube-330h-soniox.

sourceHugging Faceotherupdated 5d agoView on Hugging Face
0likes27downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ErfanRou/persian-youtube-330h-soniox · CoolFace