MadLook/arabic-whisper-multidialect-processed-small
Arabic Whisper Multi-Dialect - Processed (Small) Dataset Description This is a preprocessed version of the Arabic multi-dialect speech dataset, ready for fine-tuning OpenAI's Whisper models. The dataset contains audio features extracted and formatted specifically for Whisper training. Size: 40% subset of the full arabic-whisper-multidialect dataset Total Examples: 43,091 samples Format: Pre-computed Whisper input features (mel spectrograms) and tokenized labels… See the full description on the dataset page: https://huggingface.co/datasets/MadLook/arabic-whisper-multidialect-processed-small.
114
Update README.md
Update README.md
Upload dataset (part 00001-of-00002)
Upload dataset (part 00000-of-00002)
initial commit
