CoolFace
Datasetpublic

community-datasets/tashkeela

Dataset Card for Tashkeela Dataset Summary It contains 75 million of fully vocalized words mainly 97 books from classical and modern Arabic language. Supported Tasks and Leaderboards [More Information Needed] Languages The dataset is based on Arabic. Dataset Structure Data Instances {'book':… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/tashkeela.

sourceHugging Facegpl-2.0updated 2y agoView on Hugging Face
6likes214downloads
13 commits on main
86a22712y ago

Convert dataset to Parquet (#2)

albertvillanova
dd881f53y ago

Delete legacy JSON metadata (#1)

albertvillanova
8c3a3884y ago

add dataset_info in dataset metadata

lhoestq
459407a4y ago

remove dummmy data

mariosasko
e7d8ca74y ago

fix task_ids

lhoestq
19813314y ago

Align more metadata with other repo types (models,spaces) (#4607)

julien-c
f68af1c4y ago

Remove config names as yaml keys (#4367)

lhoestq
19679544y ago

Update datasets task tags to align tags with models (#4067)

lhoestq
2b331d95y ago

Update files from the datasets library (from 1.18.0)

system
6159a225y ago

Update files from the datasets library (from 1.7.0)

system
9ec240f5y ago

Update files from the datasets library (from 1.6.0)

system
16c5c205y ago

Update files from the datasets library (from 1.3.0)

system
bcad3d55y ago

Update files from the datasets library (from 1.2.0)

system