CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01FormosanBank /ePark_wen_hua_pian_cultural_section FormosanBank publication status This audio is associated with XML published in the public FormosanBank corpus and uses the same license recorded in that XML: CC BY-NC-SA 4.0. View the published XML. Publication approval is recorded on the corresponding FormosanBank Basecamp card. FormosanBank/ePark_wen_hua_pian_cultural_section Commercial AI Use is prohibited without prior written permission. See the FormosanBank Terms of Use and AI Use Addendum. This is a… See the full description on the dataset page: https://huggingface.co/datasets/FormosanBank/ePark_wen_hua_pian_cultural_section.audioautomatic-speech-recognition10K<n<100K0 likes294 downloads2mo agoHugging Face02pourmand1376 /asr-farsi-youtube-chunked-30-seconds How To Use from datasets import load_dataset train = load_dataset('pourmand1376/asr-farsi-youtube-chunked-30-seconds', split='train+val') test =load_dataset('pourmand1376/asr-farsi-youtube-chunked-30-seconds', split='test') +300 Hours ASR dataset generated from this kaggle dataset audioautomatic-speech-recognition10K<n<100K10 likes259 downloads3y agoHugging Face03ivangtorre /second_americas_nlp_2022 Second AmericasNLP 2022 Dataset Summary This dataset contains the speech data released as part of the Second Workshop on Natural Language Processing for Indigenous Languages of the Americas (AmericasNLP 2022). It provides audio recordings and corresponding transcriptions for several Indigenous languages of the Americas together with Spanish, supporting research on multilingual and low-resource Automatic Speech Recognition (ASR). This Hugging Face version has been… See the full description on the dataset page: https://huggingface.co/datasets/ivangtorre/second_americas_nlp_2022.audioautomatic-speech-recognition1K<n<10K0 likes234 downloads2mo agoHugging Face04Kppwdfgu1 /yoruba-second-sbpn-demucs-20260826gated yoruba-second-sbpn-demucs-20260826 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-second-sbpn-demucs-20260826.audioautomatic-speech-recognition1K<n<10K0 likes33 downloads28d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.