CoolFace
Datasetpublic

MadLook/arabic-whisper-multidialect-processed-small

Arabic Whisper Multi-Dialect - Processed (Small) Dataset Description This is a preprocessed version of the Arabic multi-dialect speech dataset, ready for fine-tuning OpenAI's Whisper models. The dataset contains audio features extracted and formatted specifically for Whisper training. Size: 40% subset of the full arabic-whisper-multidialect dataset Total Examples: 43,091 samples Format: Pre-computed Whisper input features (mel spectrograms) and tokenized labels… See the full description on the dataset page: https://huggingface.co/datasets/MadLook/arabic-whisper-multidialect-processed-small.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
1likes14downloads
5 commits on main
42179cc10mo ago

Update README.md

MadLook
74c461810mo ago

Update README.md

MadLook
7cd28d410mo ago

Upload dataset (part 00001-of-00002)

MadLook
027793110mo ago

Upload dataset (part 00000-of-00002)

MadLook
c56e57410mo ago

initial commit

MadLook