CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01datapointai /text-to-speech-human-preferences-315kgated Text-to-speech human preferences: 315K votes across 15 models This gated dataset contains the evaluation record behind Datapoint Audio Bench: 315,000 eligible pairwise votes comparing 15 text-to-speech models in a complete round-robin over 300 English prompts. The prompt set covers eight practical voice-agent categories, and every generated sample is included as a typed audio record. The source evaluation collected 357,651 completed responses. The published benchmark excluded… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-to-speech-human-preferences-315k.audiotext-to-speech100K<n<1M38 likes571 downloads23d agoHugging Face02Appenlimited /700h-tr-turkish-text-to-speechaudioautomatic-speech-recognition1K<n<10K17 likes507 downloads1y agoHugging Face03danielrosehill /Speech-To-Text-System-Prompts-2 Speech To Text System Prompt Library This repository provides a collection of system prompts designed to transform and refine text captured using speech-to-text technologies. By passing STT outputs through large language models with these specialized prompts, you can achieve cleaner, more structured, and purpose-specific text formats. 📋 The Idea Here is the basic implementation. I don't pretend that this is the stuff of high AI engineering. But it does create quite… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/Speech-To-Text-System-Prompts-2.imagen<1K1 likes306 downloads1y agoHugging Face04X-lord /Dataset-Text-To-Speech-Indonesia 🎵 Dataset Audio Bahasa Indonesia Dataset audio berkualitas tinggi untuk Text-to-Speech (TTS) bahasa Indonesia. Dibuat oleh : Muhammad Arief, S.Kom.Universitas Muhammadiyah SorongTeknik Informatika 2020 📊 Spesifikasi Teknis Parameter Nilai Satuan Total Durasi 16.38 jam Jumlah Segmen 4531 file Durasi Rata-rata 13.01 detik Sample Rate KHz 22 kHz Sample Rate Hz 22000 Hz Bit Depth PCM_16 PCM Format wav Lossless 🔄 Urutan Pengolahan… See the full description on the dataset page: https://huggingface.co/datasets/X-lord/Dataset-Text-To-Speech-Indonesia.audiotext-to-speech1K<n<10K3 likes189 downloads8mo agoHugging Face05adiren7 /darija_speech_to_textaudioautomatic-speech-recognition10K<n<100K13 likes180 downloads2y agoHugging Face06pujanpaudel /nepali_speech_to_text Nepali Speech-to-Text Dataset This repository contains a dataset for Automatic Speech Recognition (ASR) in the Nepali language. The dataset is designed for supervised learning tasks and includes audio files along with their corresponding transcriptions. The audio samples have been collected from various open-source platforms and other publicly available sources on the internet. Each audio file has an average length of 15 seconds and has been converted into a consistent WAV format… See the full description on the dataset page: https://huggingface.co/datasets/pujanpaudel/nepali_speech_to_text.audioautomatic-speech-recognition1K<n<10K1 likes154 downloads2y agoHugging Face07phuhuyqhqb /speech-to-textaudio10K<n<100K0 likes144 downloads3mo agoHugging Face08Tamazight-NLP /Tamazight-Speech-to-Arabic-Text Tamazight-Arabic Speech Recognition Dataset This is the Tamazight-NLP organization-hosted version of the Tamazight-Arabic Speech Recognition Dataset. This dataset contains ~15.5 hours of Tamazight (Tachelhit dialect) speech paired with Arabic transcriptions, designed for automatic speech recognition (ASR) and speech-to-text translation tasks. Dataset Details Total Examples: 20,344 audio segments Training Set: 18,309 examples (~8.9GB) Test Set: 2,035 examples (~992MB)… See the full description on the dataset page: https://huggingface.co/datasets/Tamazight-NLP/Tamazight-Speech-to-Arabic-Text.audiotranslation10K<n<100K8 likes137 downloads1y agoHugging Face09the-vedantic-coder /text-to-speech-en-IN-checkpoint0 likes135 downloads9mo agoHugging Face10andrewatef /Arabic-Text-to-Speechaudio10K<n<100K3 likes131 downloads1y agoHugging Face11baohuynhbk14 /vietnamese-speech-to-text-preprocessed-whisper-large-v31K<n<10K0 likes128 downloads3y agoHugging Face12lubobill1990 /speech_to_text_yixing_dialectaudio10K<n<100K0 likes115 downloads1y agoHugging Face13baohuynhbk14 /vietnamese-speech-to-text-preprocessed-whisper-medium1K<n<10K1 likes111 downloads3y agoHugging Face14BrunoHays /darija-speech-to-text Speech To Text Darija dataset Reupload of adiren7/darija_speech_to_text audioautomatic-speech-recognition1K<n<10K6 likes83 downloads2y agoHugging Face15EMINES /Tamazight-Speech-to-Arabic-Text Tamazight-Arabic Speech Recognition Dataset Overview This is the EMINES organization-hosted version of the Tamazight-Arabic Speech Recognition Dataset, synchronized with the original dataset. It contains ~15.5 hours of Tamazight speech (Tachelhit dialect) paired with Arabic transcriptions, designed for developing ASR and translation systems. Quick Start from datasets import load_dataset # Load the dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/EMINES/Tamazight-Speech-to-Arabic-Text.audioautomatic-speech-recognition10K<n<100K6 likes64 downloads2y agoHugging Face16crtvai /arabic_speech_to_text_20241219_205753_x4mhwqaudio1K<n<10K0 likes63 downloads2y agoHugging Face17crtvai /arabic_speech_to_text_20241219_203218_lp9vcvaudio1K<n<10K0 likes61 downloads2y agoHugging Face18crtvai /arabic_speech_to_text_20241218_144737_gkopimaudion<1K2 likes56 downloads2y agoHugging Face19Charif-Ayfarah /Afar-language-text-to-speech-TTS Usage This dataset is designed to support the development of Text-to-Speech (TTS) systems for the Afar language. It can be integrated into web applications, mobile apps, desktop software, or other platforms that require natural-sounding Afar voice synthesis or accurate spoken language recognition. For applications involving virtual avatars or voice personas, the following culturally appropriate voice names are recommended: Female Voices: Emeli, Hanaawi, Kareera, Laysani, Kulsuma… See the full description on the dataset page: https://huggingface.co/datasets/Charif-Ayfarah/Afar-language-text-to-speech-TTS.text-to-speech1 likes56 downloads10mo agoHugging Face20adiren7 /darija_to_french_speech_to_textaudion<1K7 likes47 downloads2y agoHugging Face21crtvai /arabic_speech_to_text_20241218_174023_sd5hyyaudio1K<n<10K0 likes42 downloads2y agoHugging Face22amitpant7 /nepali-speech-to-textHere's a README draft for your Hugging Face dataset: Nepali Speech-to-Text Dataset This dataset contains high-quality speech samples in Nepali, originally from OpenSLR SLR43 and Mozilla's Common Voice dataset. It has been cleaned and processed for Automatic Speech Recognition (ASR) tasks. The dataset consists of approximately 3,000 audio samples, each around 30 seconds long, compiled for use in training and testing ASR models. Dataset Details Number of samples:… See the full description on the dataset page: https://huggingface.co/datasets/amitpant7/nepali-speech-to-text.audio1K<n<10K0 likes36 downloads2y agoHugging Face23riyaSingh310504 /TextToSpeechMedConvoaudion<1K1 likes34 downloads4mo agoHugging Face2434data /nepali-speech-to-textaudion<1K0 likes34 downloads2mo agoHugging Face25crtvai /arabic_speech_to_text_20241224_135643_72lw7raudio1K<n<10K0 likes33 downloads2y agoHugging Face26harshal-07 /speech_to_texttabularn<1K0 likes29 downloads3y agoHugging Face27crtvai /arabic_speech_to_text_20241224_134331_eiicpwaudio1K<n<10K1 likes29 downloads2y agoHugging Face28crtvai /arabic_speech_to_text_20241218_173148_7rkwsyaudio1K<n<10K0 likes27 downloads2y agoHugging Face29crtvai /arabic_speech_to_text_20241222_184338_sngqcgaudion<1K0 likes27 downloads2y agoHugging Face30LeVy4 /speech-to-textaudiotext-to-speechn<1K0 likes25 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.