gfbati/Ten2Zero
This dataset contains the following: 1- A balanced audio dataset of spoken Arabic digits from ten to zero in wav form (located at the "Dataset" folder); 2- A balanced image dataset of spoken Arabic digits from ten to zero in png form (located at the "Dataset" folder); 3- Tabular data generated using deep learning (SqueezeNet and Inception v3) from the spectrograms of the audio files; 4- Orange Data Mining workflows (".ows" files) used in processing this dataset. Please cite the following… See the full description on the dataset page: https://huggingface.co/datasets/gfbati/Ten2Zero.
171
1version https://git-lfs.github.com/spec/v12oid sha256:e720fc7faf6a5a26d01f527c4e294c6f44d3b18f1a9ca1d73fb3635e7b8956db3size 175035494 