CoolFace
Datasetpublic

Peacockery/farsi-asr-wer35-fastconformer

Farsi ASR WER35 FastConformer This dataset contains Farsi/Farsi ASR utterances curated with NVIDIA NeMo Curator using nvidia/stt_fa_fastconformer_hybrid_large. The uploaded training data is stored as WebDataset TAR shards because Hugging Face recommends WebDataset archives for large-scale audio datasets. The local artifact also includes a NeMo ASR manifest at manifests/train_manifest.jsonl. Curation Language: Farsi/Farsi (fa) Audio: FLAC, 16 kHz mono Source run:… See the full description on the dataset page: https://huggingface.co/datasets/Peacockery/farsi-asr-wer35-fastconformer.

sourceHugging Faceupdated 4mo agoView on Hugging Face
0likes6downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
Peacockery/farsi-asr-wer35-fastconformer · CoolFace