CoolFace
Datasetpublic

Reza2kn/persian-asr-text-2.69M-deduped

🗂️ persian-asr-text-2.69M-deduped English + فارسی · Part of Shenava 1.0 · Project hub · SLT paper submission 🌟 At a glance | معرفی سریع English فارسی 🎯 Purpose Deduplicated Persian ASR text dataset used by the training stack. پیکرهٔ متنی فارسیِ حذف‌تکرارشده برای ساخت واژگان، مدل‌سازی زبانی و پشتیبانی از آموزش ASR. 🧩 Role Persian text and linguistic asset مصنوع متنی و زبانی فارسی 📦 Snapshot 4 files; approximately 109.64 MB 4 فایل؛ حدود 109.64… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-asr-text-2.69M-deduped.

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
0likes69downloads
6 commits on main
364f6e72mo ago

Add extensive bilingual English-Persian Shenava-1 card

Reza2kn
9a68fcb2mo ago

Remove stale non-Apache license text

Reza2kn
19760162mo ago

Set Shenava artifact license to Apache-2.0

Reza2kn
feb84025mo ago

Upload README.md with huggingface_hub

Reza2kn
cb72d705mo ago

Upload data/data.parquet with huggingface_hub

Reza2kn
1dfe8125mo ago

initial commit

Reza2kn