CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01KTH /hungarian-single-speaker-tts Dataset Card for CSS10 Hungarian: Single Speaker Speech Dataset Dataset Summary The corpus consists of a single speaker, with 4515 segments extracted from a single LibriVox audiobook. Supported Tasks and Leaderboards [Needs More Information] Languages The audio is in Hungarian. Dataset Structure [Needs More Information] Data Instances [Needs More Information] Data Fields [Needs More Information] Data Splits… See the full description on the dataset page: https://huggingface.co/datasets/KTH/hungarian-single-speaker-tts.audiotext-to-speech1K<n<10K14 likes187 downloads4y agoHugging Face02datadriven-company /TTS-Hungarian TTS-Hungarian A large-scale, high-quality Hungarian speech dataset for text-to-speech and automatic speech recognition. Data Source Derived from MEK (Magyar Elektronikus Könyvtár) — Hungarian audiobooks. Dataset Statistics Metric Value Total samples 253,116 Total duration 702 hours Unique speakers 100 Average duration 10.0 seconds Average DNSMOS 3.68 Features Field Type Description __key__ string Unique sample… See the full description on the dataset page: https://huggingface.co/datasets/datadriven-company/TTS-Hungarian.audiotext-to-speech100K<n<1M1 likes179 downloads7mo agoHugging Face03shunyalabs /hungarian-speech-datasetaudio1K<n<10K0 likes60 downloads1y agoHugging Face04Speech-data /Hungarian-Speech-Dataset 🎧 Hungarian Speech Dataset The Hungarian Speech Dataset is a high-quality speech audio dataset designed to support advanced AI systems that depend on diverse audio data and reliable voice data for multilingual model training. It comprises 169 hours of recordings across 743 files, provided in MP3 and WAV formats, with a total size of 134 MB. This structured audio dataset ensures balanced speaker representation, featuring 46% female and 54% male speakers, and an age distribution… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Hungarian-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes32 downloads6mo agoHugging Face05Thomcles /YodaLingua-Hungariangated YodaLingua-Hungarian YodaLingua is a high-quality speech dataset designed for training text-to-speech (TTS) systems, ASR models, and any application requiring clean, well-aligned audio–text pairs.This release contains the Hungarian portion of the multilingual YodaLingua collection. 🧾 Dataset Overview Property Value Total clips 80,740 audio–transcription pairs Total duration 206 hours Speakers 1,856 distinct speakers Audio format MP3 • mono • 24 kHz •… See the full description on the dataset page: https://huggingface.co/datasets/Thomcles/YodaLingua-Hungarian.audiotext-to-speech10K<n<100K0 likes20 downloads5mo agoHugging Face06eleferrand /Voxpopuli_Hungarian Dataset Card for "Voxpopuli_Hungarian" More Information needed audio1K<n<10K0 likes17 downloads9mo agoHugging Face07Hungarians /Samplesaudion<1K0 likes8 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.