CoolFace
Datasetpublic

oddadmix/arabic-audio-collection-syrian-podcast

Syrian Postcast Arabic Speech Dataset Dataset Summary The Syrian Postcast Arabic Speech Dataset is a large-scale, first-of-its-kind Arabic speech corpus containing approximately 116 hours of speech recordings and corresponding transcripts. What distinguishes this dataset as a pioneering resource in Arabic language technology is its comprehensive inclusion of rich non-verbal transcriptions. Alongside the spoken Arabic text, the transcripts meticulously capture… See the full description on the dataset page: https://huggingface.co/datasets/oddadmix/arabic-audio-collection-syrian-podcast.

sourceHugging Faceotherupdated 3mo agoView on Hugging Face
2likes138downloads
7 commits on main
4f088ed3mo ago

Update README.md

oddadmix
e97756b3mo ago

Update README.md

oddadmix
6b9dbcd4mo ago

Update README.md

oddadmix
3c1d3f94mo ago

Update README.md

oddadmix
2b801f14mo ago

Update README.md

oddadmix
3412c144mo ago

Upload dataset

oddadmix
f3b24544mo ago

initial commit

oddadmix