CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Hani89 /Synthetic-Medical-Speech-Dataset Synthetic Medical Speech Dataset Overview Synthetic Medical Speech Dataset is a synthetic dataset of audio–text pairs designed for developing and evaluating automatic speech recognition (ASR) models in the medical domain.The corpus contains thousands of short audio clips generated from medically relevant text using a text-to-speech (TTS) system.Each clip is paired with its corresponding transcript.Because all content is synthetically produced, the dataset does not contain… See the full description on the dataset page: https://huggingface.co/datasets/Hani89/Synthetic-Medical-Speech-Dataset.audioautomatic-speech-recognition10K<n<100K4 likes248 downloads11mo agoHugging Face02VladS159 /romanian_speech_dataset_with_15_percent_6_speakers_synthetic_dataaudio10K<n<100K0 likes188 downloads7mo agoHugging Face03mnozxe /nazrah-synthetic-datasetaudion<1K0 likes182 downloads3mo agoHugging Face04CLEAR-Global /Hausa-Synthetic-ASR-Dataset-XTTSgatedSynthetic Hausa ASR dataset generated using a fine-tuned version of the XTTS-v2 model. Sample rate: 24kHz. Total duration: 574 hours. audioautomatic-speech-recognition100K<n<1M1 likes146 downloads1y agoHugging Face05Wi-Fi /korean-full-duplex-synthetic-dataset-preview Korean Full-Duplex Synthetic Dataset Preview Overview Public preview of a Korean full-duplex synthetic speech dataset. This repository contains 100 conversations sampled from a corpus of 89,273 conversations (2,000.5 hours); it does not publish the full corpus audio. Preview contents 100 conversation WAV files data/representative.jsonl 24 kHz, mono, 16-bit PCM Events: normal, barge_in, backchannel, cutoff_by_user Annotation format… See the full description on the dataset page: https://huggingface.co/datasets/Wi-Fi/korean-full-duplex-synthetic-dataset-preview.audioautomatic-speech-recognitionn<1K1 likes138 downloads1mo agoHugging Face06Cossale /synthetic-gujarati-tts-datasetaudio1K<n<10K0 likes131 downloads2y agoHugging Face07Anilosan15 /Synthetic_Turkish_TTS_Data Synthetic Turkish TTS Data This dataset was created by generating synthetic Turkish text across multiple speech scenarios. The text was produced in the following domains: finance_master, cs_master, parcel_delivery, ecommerce, telecom, isp_support, technical_support, subscription, insurance, health_appointments, public_services, education_registration, and daily_speech. These synthetic texts were then synthesized with a high-quality Turkish TTS model. The dataset is intended to be… See the full description on the dataset page: https://huggingface.co/datasets/Anilosan15/Synthetic_Turkish_TTS_Data.audiotext-to-speech10K<n<100K6 likes116 downloads5mo agoHugging Face08VladS159 /romanian_speech_dataset_with_20_percent_4_speakers_synthetic_dataaudio10K<n<100K0 likes111 downloads6mo agoHugging Face09uncleMehrzad /synthetic-speaker-diarization-dataset-fa-large-3000audioaudio-classification1K<n<10K3 likes108 downloads1y agoHugging Face10DatarrX /burmese-synthetic-speech-corpus Burmese Synthetic Speech Corpus (DatarrX/burmese-synthetic-speech-corpus) Overview The Burmese Synthetic Speech Corpus is a high-fidelity, manually curated audio dataset specifically designed to advance Text-to-Speech (TTS) systems, speech recognition, and other audio-driven Machine Learning tasks for the Burmese (Myanmar) language. Created by DatarrX, this dataset bridges the gap in low-resource speech technologies by providing highly natural, native-sounding… See the full description on the dataset page: https://huggingface.co/datasets/DatarrX/burmese-synthetic-speech-corpus.audiotext-to-speech1K<n<10K7 likes107 downloads4mo agoHugging Face11shreyaskal3 /synthetic-speaker-diarization-dataset-hindiaudion<1K0 likes69 downloads2y agoHugging Face12whitneyten /synthetic-data-japanaudio1K<n<10K0 likes68 downloads1y agoHugging Face13bismarck91 /fr_en_synthetic_audio_datasetaudio10K<n<100K0 likes60 downloads1y agoHugging Face14whitneyten /synthetic-data-indonesia_testaudion<1K0 likes52 downloads1y agoHugging Face15diarizers-community /synthetic-speaker-diarization-datasetaudio1K<n<10K2 likes47 downloads2y agoHugging Face16whitneyten /synthetic-data-indonesiaaudio1K<n<10K0 likes45 downloads1y agoHugging Face17Samyak29 /synthetic-speaker-diarization-dataset-hindi-largeaudion<1K1 likes43 downloads2y agoHugging Face18VladS159 /romanian_speech_dataset_with_40_percent_8_speakers_synthetic_dataaudio10K<n<100K0 likes43 downloads6mo agoHugging Face19whitneyten /synthetic-data-indonesia_2_4audion<1K0 likes41 downloads1y agoHugging Face20kamilakesbi /synthetic_dataset_jpn_2_more_speakersaudio1K<n<10K0 likes40 downloads2y agoHugging Face21whitneyten /synthetic-data-indonesia_2_4_updatedaudio1K<n<10K0 likes39 downloads1y agoHugging Face22Taklaxbr /Synthetic_Turkish_TTS_Data Not: Bu veri setinin dokümantasyonu Türk yapay zeka topluluğuna katkı sağlamak amacıyla VeriPazarı tarafından Türkçeye çevrilmiştir. Orijinal veri seti Anilosan15 tarafından geliştirilmiş olup, VeriPazarı tarafından Türk AI ekosistemi için arşivlenmiştir. 🔗 Orijinal Kaynak: Anilosan15/Synthetic_Turkish_TTS_Data 🔗 Derleyen Platform: VeriPazarı Sentetik Türkçe TTS Veri Seti (Synthetic Turkish TTS Data) Bu veri seti, çoklu konuşma senaryoları üzerinden sentetik Türkçe metinler… See the full description on the dataset page: https://huggingface.co/datasets/Taklaxbr/Synthetic_Turkish_TTS_Data.audiotext-to-speech10K<n<100K1 likes39 downloads3mo agoHugging Face23Tuyentd /Emotion_conversation_synthetic_datasetaudio1K<n<10K0 likes38 downloads1y agoHugging Face24kamilakesbi /synthetic_dataset_jpnaudio1K<n<10K0 likes35 downloads2y agoHugging Face25kamilakesbi /synthetic_dataset_jpn_volumesaudio1K<n<10K0 likes28 downloads2y agoHugging Face26maheshbabu9199 /synthetic-speaker-diarization-datasetaudion<1K0 likes27 downloads2y agoHugging Face27CLEAR-Global /Chichewa-Synthetic-ASR-DatasetgatedSynthetic Chichewa ASR dataset generated using a fine-tuned version of the YourTTS model. Sample rate: 24kHz. Total duration: 550 hours. audioautomatic-speech-recognition100K<n<1M2 likes27 downloads1y agoHugging Face28uncleMehrzad /synthetic-speaker-diarization-dataset-fasynthetic dataset generated from persian common voice. audion<1K1 likes24 downloads1y agoHugging Face29Tuyentd /Emotion_conversation_synthetic_dataset_shortaudio1K<n<10K0 likes24 downloads11mo agoHugging Face30MohamedGomaa30 /Synthetic-Egy-Speech-Dataset Synthetic Egyptian Speech Dataset A curated dataset of 1000 Egyptian Arabic speech samples — the best audio selected across 4 TTS models for each prompt, with transcription text and quality metadata. Dataset Description Each entry contains: id: Unique prompt identifier (e.g., egy_0001) text: Egyptian Arabic transcription text audio_path: Path to the best-selected .wav audio file model: TTS model that produced the best audio (lahgtna, chatterbox_egyptian, egtts_v01, or… See the full description on the dataset page: https://huggingface.co/datasets/MohamedGomaa30/Synthetic-Egy-Speech-Dataset.audiotext-to-speech1K<n<10K2 likes24 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.