CoolFace
Datasetpublic

Reza2kn/shenava-koochik-number-stress-36k

Shenava Koochik Number Stress 36K This repository contains the 36,000-row Persian number-stress text curriculum and an audio-backed subset of 6,669 short clips generated with Gemini TTS. The Dataset Viewer default configuration is the 6,669-clip audio subset. It exposes audio, text, category, and duration_s; the metadata links each row to its WAV using a relative file_name such as audio/num-000000.wav. The complete text-only source remains available at data/train.jsonl (36,000… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/shenava-koochik-number-stress-36k.

sourceHugging Faceupdated 19d agoView on Hugging Face
1likes264downloads
Dataset Card

Shenava Koochik Number Stress 36K

This repository contains the 36,000-row Persian number-stress text curriculum and an audio-backed subset of 6,669 short clips generated with Gemini TTS.

The Dataset Viewer default configuration is the 6,669-clip audio subset. It exposes audio, text, category, and duration_s; the metadata links each row to its WAV using a relative file_name such as audio/num-000000.wav.

The complete text-only source remains available at data/train.jsonl (36,000 rows, with text and category fields).

The text utterances are constrained to 2–8 whitespace-separated words and contain no Arabic/Persian digits or Latin letters. Categories cover clock time, dates, cardinals, ordinals, arithmetic, money, percentages, fractions, ranges, measures, and sequence IDs.

The text contexts were requested from z-ai/glm-5.3-flash using short number-centric in-context examples. Rows that failed hard constraints were completed with deterministic short templates and retained in the same category distribution.