CoolFace
Datasetpublic

taqwa92/cm.trial

Dataset Card for Common Voice Corpus 11.0 Dataset Summary The Common Voice dataset consists of a unique MP3 and corresponding text file. Many of the 24210 recorded hours in the dataset also include demographic metadata like age, sex, and accent that can help improve the accuracy of speech recognition engines. The dataset currently consists of 16413 validated hours in 100 languages, but more voices and languages are always added. Take a look at the Languages… See the full description on the dataset page: https://huggingface.co/datasets/taqwa92/cm.trial.

sourceHugging Facecc0-1.0updated 4y agoView on Hugging Face
0likes54downloads
6 commits on main
c9587564y ago

Upload README.md

taqwa92
d022a6b4y ago

Upload 4 files

taqwa92
ad660724y ago

Upload common_voice_11_0.py

taqwa92
55f41954y ago

Upload transcript/ar with huggingface_hub

taqwa92
aa3a04f4y ago

Upload audio/ar/dev with huggingface_hub

taqwa92
b5855634y ago

initial commit

taqwa92