mediaspeech
mediaspeech-with-cv-tr
Dataset Card for "mediaspeech-with-cv-tr"
More Information needed
MediaSpeech
MediaSpeech
MediaSpeech is a dataset of Arabic, French, Spanish, and Turkish media speech built with the purpose of testing Automated Speech Recognition (ASR) systems performance. The dataset contains 10 hours of speech for each language provided.
The dataset consists of short speech segments automatically extracted from media videos available on YouTube and manually transcribed, with some pre-processing and post-processing.
Baseline models and WAV version of the dataset can be… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/MediaSpeech.stt-mediaspeech-test
MediaSpeech — French test split
Split test de MediaSpeech (français) — extraits courts de médias (radio,
TV, podcasts) collectés par MTS AI. Empaqueté en Parquet shardé avec audio
FLAC embarqué.
Usage principal : benchmark ASR français (WER / CER) sur parole de
diffusion (broadcast / médias).
Contenu
2498 utterances (segments ~10 s)
Audio : FLAC 16 kHz mono PCM_16
Langue : français (fr)
Licence : CC-BY-4.0 (héritée de MediaSpeech / OpenSLR 108)
Durée totale : 10.00 h… See the full description on the dataset page: https://huggingface.co/datasets/ggfox00000/stt-mediaspeech-test.MediaSpeech_arIraqi_shiekh_hara_sawtarabi_MediaSpeechAll_WSP_Iraqi_shiekh_hara_sawtarabi_MediaSpeech
