CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01LocalDoc /azerbaijani_asr Azerbaijani ASR Dataset Dataset Description This dataset contains Azerbaijani speech data for Automatic Speech Recognition (ASR) tasks. Dataset Summary Language: Azerbaijani (az) Task: Automatic Speech Recognition Total Duration: ~328 hours Total Samples: ~345,643 audio-text pairs Audio Format: WAV, 16kHz sampling rate License: CC-BY-4.0 Dataset Structure Each audio segment is specially numbered so that you can merge them if you… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/azerbaijani_asr.audioautomatic-speech-recognition100K<n<1M4 likes754 downloads2mo agoHugging Face02TonyFANgr /localaiiaasr Emergency Room FAQ ASR Evaluation Set Small audio evaluation set for testing automatic speech recognition (ASR) on emergency-room FAQ questions, recorded in Taiwanese Hokkien (nan), Mandarin (cmn), and code-switched Hokkien/Mandarin (mixed). Most questions are spoken in both Hokkien and Mandarin, so most transcriptions have a matching pair of audio clips; a small subset also has a mixed-language clip. The questions come from eval/FAQ_dataset in the local_aiia project, grouped… See the full description on the dataset page: https://huggingface.co/datasets/TonyFANgr/localaiiaasr.audioautomatic-speech-recognitionn<1K0 likes104 downloads1mo agoHugging Face03LocalDoc /fleurs-azerbaijani-asr FLEURS Azerbaijani ASR Benchmark Azerbaijani (az_az) subset of FLEURS, reformatted for ASR benchmarking and fine-tuning. Source Based on FLEURS dataset by Google (Conneau et al., 2022). Licensed under CC-BY-4.0. Structure Split Samples Duration train 2656 9.28h dev 400 1.35h test 921 3.23h Fields audio — 16kHz mono WAV sentence — transcription (original casing and punctuation) sentence_normalized — normalized (lowercase, no… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/fleurs-azerbaijani-asr.audioautomatic-speech-recognition1K<n<10K0 likes93 downloads6mo agoHugging Face04niloycste68 /Bangali_local_dialect_ASR_HF_Dataset BanglaMix — Code-Switching ASR in Bangladeshi Regional Dialects BanglaMix is a speech dataset for Automatic Speech Recognition (ASR) on dialectal Bangladeshi Bengali mixed with English (code-switching). It covers 15 regional dialects and the natural Bengali–English code-switching common in informal Bangladeshi speech — a setting not covered by existing Bengali corpora, which address either dialects or code-switching, never both. Clips 41,499 transcribed audio clips… See the full description on the dataset page: https://huggingface.co/datasets/niloycste68/Bangali_local_dialect_ASR_HF_Dataset.audioautomatic-speech-recognition10K<n<100K0 likes47 downloads4mo agoHugging Face05Jarbas /localingua_pt-brtranscriptions unverified! known to contain mistakes/noise audioautomatic-speech-recognitionn<1K1 likes27 downloads2y agoHugging Face06TigreGotico /locallingua_ptRecordings from Portugal downloaded from https://localingual.com audioautomatic-speech-recognitionn<1K0 likes26 downloads2y agoHugging Face07Jarbas /localingua_africa_ptaudioautomatic-speech-recognitionn<1K1 likes4 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.