CoolFace
14 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Hani89 /Synthetic-Medical-Speech-Dataset Synthetic Medical Speech Dataset Overview Synthetic Medical Speech Dataset is a synthetic dataset of audio–text pairs designed for developing and evaluating automatic speech recognition (ASR) models in the medical domain.The corpus contains thousands of short audio clips generated from medically relevant text using a text-to-speech (TTS) system.Each clip is paired with its corresponding transcript.Because all content is synthetically produced, the dataset does not contain… See the full description on the dataset page: https://huggingface.co/datasets/Hani89/Synthetic-Medical-Speech-Dataset.audioautomatic-speech-recognition10K<n<100K4 likes248 downloads11mo agoHugging Face02CLEAR-Global /Hausa-Synthetic-ASR-Dataset-XTTSgatedSynthetic Hausa ASR dataset generated using a fine-tuned version of the XTTS-v2 model. Sample rate: 24kHz. Total duration: 574 hours. audioautomatic-speech-recognition100K<n<1M1 likes146 downloads1y agoHugging Face03Wi-Fi /korean-full-duplex-synthetic-dataset-preview Korean Full-Duplex Synthetic Dataset Preview Overview Public preview of a Korean full-duplex synthetic speech dataset. This repository contains 100 conversations sampled from a corpus of 89,273 conversations (2,000.5 hours); it does not publish the full corpus audio. Preview contents 100 conversation WAV files data/representative.jsonl 24 kHz, mono, 16-bit PCM Events: normal, barge_in, backchannel, cutoff_by_user Annotation format… See the full description on the dataset page: https://huggingface.co/datasets/Wi-Fi/korean-full-duplex-synthetic-dataset-preview.audioautomatic-speech-recognitionn<1K1 likes138 downloads1mo agoHugging Face04Anilosan15 /Synthetic_Turkish_TTS_Data Synthetic Turkish TTS Data This dataset was created by generating synthetic Turkish text across multiple speech scenarios. The text was produced in the following domains: finance_master, cs_master, parcel_delivery, ecommerce, telecom, isp_support, technical_support, subscription, insurance, health_appointments, public_services, education_registration, and daily_speech. These synthetic texts were then synthesized with a high-quality Turkish TTS model. The dataset is intended to be… See the full description on the dataset page: https://huggingface.co/datasets/Anilosan15/Synthetic_Turkish_TTS_Data.audiotext-to-speech10K<n<100K6 likes116 downloads5mo agoHugging Face05uncleMehrzad /synthetic-speaker-diarization-dataset-fa-large-3000audioaudio-classification1K<n<10K3 likes108 downloads1y agoHugging Face06DatarrX /burmese-synthetic-speech-corpus Burmese Synthetic Speech Corpus (DatarrX/burmese-synthetic-speech-corpus) Overview The Burmese Synthetic Speech Corpus is a high-fidelity, manually curated audio dataset specifically designed to advance Text-to-Speech (TTS) systems, speech recognition, and other audio-driven Machine Learning tasks for the Burmese (Myanmar) language. Created by DatarrX, this dataset bridges the gap in low-resource speech technologies by providing highly natural, native-sounding… See the full description on the dataset page: https://huggingface.co/datasets/DatarrX/burmese-synthetic-speech-corpus.audiotext-to-speech1K<n<10K7 likes107 downloads4mo agoHugging Face07Taklaxbr /Synthetic_Turkish_TTS_Data Not: Bu veri setinin dokümantasyonu Türk yapay zeka topluluğuna katkı sağlamak amacıyla VeriPazarı tarafından Türkçeye çevrilmiştir. Orijinal veri seti Anilosan15 tarafından geliştirilmiş olup, VeriPazarı tarafından Türk AI ekosistemi için arşivlenmiştir. 🔗 Orijinal Kaynak: Anilosan15/Synthetic_Turkish_TTS_Data 🔗 Derleyen Platform: VeriPazarı Sentetik Türkçe TTS Veri Seti (Synthetic Turkish TTS Data) Bu veri seti, çoklu konuşma senaryoları üzerinden sentetik Türkçe metinler… See the full description on the dataset page: https://huggingface.co/datasets/Taklaxbr/Synthetic_Turkish_TTS_Data.audiotext-to-speech10K<n<100K1 likes39 downloads3mo agoHugging Face08CLEAR-Global /Chichewa-Synthetic-ASR-DatasetgatedSynthetic Chichewa ASR dataset generated using a fine-tuned version of the YourTTS model. Sample rate: 24kHz. Total duration: 550 hours. audioautomatic-speech-recognition100K<n<1M2 likes27 downloads1y agoHugging Face09MohamedGomaa30 /Synthetic-Egy-Speech-Dataset Synthetic Egyptian Speech Dataset A curated dataset of 1000 Egyptian Arabic speech samples — the best audio selected across 4 TTS models for each prompt, with transcription text and quality metadata. Dataset Description Each entry contains: id: Unique prompt identifier (e.g., egy_0001) text: Egyptian Arabic transcription text audio_path: Path to the best-selected .wav audio file model: TTS model that produced the best audio (lahgtna, chatterbox_egyptian, egtts_v01, or… See the full description on the dataset page: https://huggingface.co/datasets/MohamedGomaa30/Synthetic-Egy-Speech-Dataset.audiotext-to-speech1K<n<10K2 likes24 downloads4mo agoHugging Face10CLEAR-Global /Luo-Synthetic-ASR-DatasetgatedSynthetic Dholuo ASR dataset generated using a fine-tuned version of the YourTTS model. Sample rate: 24kHz. Total duration: 775 hours. audioautomatic-speech-recognition100K<n<1M1 likes19 downloads1y agoHugging Face11niamhtracey1 /Synthetic-Medical-Speech-Dataset Synthetic Medical Speech Dataset Overview Synthetic Medical Speech Dataset is a synthetic dataset of audio–text pairs designed for developing and evaluating automatic speech recognition (ASR) models in the medical domain.The corpus contains thousands of short audio clips generated from medically relevant text using a text-to-speech (TTS) system.Each clip is paired with its corresponding transcript.Because all content is synthetically produced, the dataset does not contain… See the full description on the dataset page: https://huggingface.co/datasets/niamhtracey1/Synthetic-Medical-Speech-Dataset.audioautomatic-speech-recognition10K<n<100K0 likes17 downloads4mo agoHugging Face12CLEAR-Global /Hausa-Synthetic-ASR-Dataset-YourTTSgatedSynthetic Hausa ASR dataset generated using a fine-tuned version of the YourTTS model. Sample rate: 24kHz. Total duration: 993 hours. audioautomatic-speech-recognition100K<n<1M0 likes12 downloads1y agoHugging Face13LisanneH /Synthetic_Speech_Data_Project Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/LisanneH/Synthetic_Speech_Data_Project.automatic-speech-recognition0 likes6 downloads3y agoHugging Face14sajalmadan0909 /combined_synthetic_datasets_eng_hin_engandhincodemixgated Combined Synthetic Datasets (English, Hindi, Code-Mix) Public ASR training data combining YouTube podcast VAD clips, English/Hinglish podcasts, and synthetic Hinglish entity-normalization speech. Subsets Config Rows Description yt_video_transcript 4,100 Hindi-dominant YouTube podcast segments (VAD chunks) vad_english 1,160 English podcast segments vad_hindi_english 787 Hindi–English code-mixed podcast segments synthetic_voice_stt 24,459 Synthetic… See the full description on the dataset page: https://huggingface.co/datasets/sajalmadan0909/combined_synthetic_datasets_eng_hin_engandhincodemix.audioautomatic-speech-recognition10K<n<100K1 likes2 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.