CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Appenlimited /700h-tr-turkish-text-to-speechaudioautomatic-speech-recognition1K<n<10K16 likes507 downloads1y agoHugging Face02adiren7 /darija_speech_to_textaudioautomatic-speech-recognition10K<n<100K13 likes180 downloads2y agoHugging Face03pujanpaudel /nepali_speech_to_text Nepali Speech-to-Text Dataset This repository contains a dataset for Automatic Speech Recognition (ASR) in the Nepali language. The dataset is designed for supervised learning tasks and includes audio files along with their corresponding transcriptions. The audio samples have been collected from various open-source platforms and other publicly available sources on the internet. Each audio file has an average length of 15 seconds and has been converted into a consistent WAV format… See the full description on the dataset page: https://huggingface.co/datasets/pujanpaudel/nepali_speech_to_text.audioautomatic-speech-recognition1K<n<10K1 likes154 downloads2y agoHugging Face04BrunoHays /darija-speech-to-text Speech To Text Darija dataset Reupload of adiren7/darija_speech_to_text audioautomatic-speech-recognition1K<n<10K6 likes83 downloads2y agoHugging Face05EMINES /Tamazight-Speech-to-Arabic-Text Tamazight-Arabic Speech Recognition Dataset Overview This is the EMINES organization-hosted version of the Tamazight-Arabic Speech Recognition Dataset, synchronized with the original dataset. It contains ~15.5 hours of Tamazight speech (Tachelhit dialect) paired with Arabic transcriptions, designed for developing ASR and translation systems. Quick Start from datasets import load_dataset # Load the dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/EMINES/Tamazight-Speech-to-Arabic-Text.audioautomatic-speech-recognition10K<n<100K6 likes64 downloads2y agoHugging Face06Charif-Ayfarah /Afar-language-text-to-speech-TTS Usage This dataset is designed to support the development of Text-to-Speech (TTS) systems for the Afar language. It can be integrated into web applications, mobile apps, desktop software, or other platforms that require natural-sounding Afar voice synthesis or accurate spoken language recognition. For applications involving virtual avatars or voice personas, the following culturally appropriate voice names are recommended: Female Voices: Emeli, Hanaawi, Kareera, Laysani, Kulsuma… See the full description on the dataset page: https://huggingface.co/datasets/Charif-Ayfarah/Afar-language-text-to-speech-TTS.text-to-speech1 likes56 downloads10mo agoHugging Face07pavi1561 /Sinhala_speech_to_textaudioautomatic-speech-recognition10K<n<100K0 likes13 downloads2y agoHugging Face08Omarrran /100_shruk-speech_to_Text__ASR_datasetgated 100 Speech-to-Text / ASR Shruk Dataset Dataset Description Traditional Kashmiri poetic verses (Shruks) paired with their Kashmiri script transcriptions. Designed for training and evaluating Automatic Speech Recognition (ASR) / Speech-to-Text (STT) models for the Kashmiri language. Dataset Structure Field Type Description audio Audio WAV recording of the Shruk transcription string Kashmiri script transcription shruk_number int Original shruk… See the full description on the dataset page: https://huggingface.co/datasets/Omarrran/100_shruk-speech_to_Text__ASR_dataset.audioautomatic-speech-recognitionn<1K0 likes7 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.