CoolFace
13 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01whatyoudoing /michii-swedish-s2-dataset Michii Swedish Speech Dataset (Fish Speech Ready) High-fidelity Swedish speech dataset generated via GPU faster-whisper-large-v3 with brand dictionary cleaning. Total Audio Clips: 535 Audio Spec: 22,050 Hz, Mono, 16-bit PCM WAV Structure: metadata.csv and .lab transcript files. audio0 likes314 downloads2mo agoHugging Face02datadriven-company /TTS-Swedish TTS-Swedish A high-quality Swedish speech dataset for text-to-speech and automatic speech recognition. Data Source Derived from LibriVox — Swedish audiobooks. Dataset Statistics Metric Value Total samples 14,535 Total duration 40 hours Unique speakers 9 Average duration 10.0 seconds Average DNSMOS 3.69 Gender Distribution Gender Samples Hours Male 11,219 31.2 Female 3,316 9.3 Features Field… See the full description on the dataset page: https://huggingface.co/datasets/datadriven-company/TTS-Swedish.audiotext-to-speech10K<n<100K0 likes132 downloads7mo agoHugging Face03yasakoko /sweet-potatoaudion<1K0 likes79 downloads3y agoHugging Face04felixmr1 /librivox-tts-swedish LibriVox TTS Swedish (5-minute chunks) Long-form variant of datadriven-company/TTS-Swedish: per-speaker audio concatenated in source order into ~5 minute chunks (16 kHz mono), for long-context TTS / ASR training. Loading from datasets import load_dataset ds = load_dataset("felixmr1/librivox-tts-swedish", split="train") # no held-out split is provided — slice it yourself, e.g. ds.train_test_split(test_size=0.05) Columns Field Type Description… See the full description on the dataset page: https://huggingface.co/datasets/felixmr1/librivox-tts-swedish.audiotext-to-speechn<1K1 likes78 downloads5mo agoHugging Face05cubbk /audio_swedish_2_dataset_cleanedaudio1K<n<10K0 likes56 downloads1y agoHugging Face06sweetcocoa /pop2piano_ciaudion<1K1 likes38 downloads3y agoHugging Face07zorrodash /mashup-xs-sweater-weatheraudion<1K0 likes37 downloads8d agoHugging Face08Speech-data /Swedish-Speech-Dataset 🎧 Swedish Speech Dataset The Swedish Speech Dataset is a high-quality speech audio dataset designed to support advanced AI and machine learning workflows with structured and diverse audio data. It contains 162 hours of voice recordings distributed across 558 files, stored in MP3 and WAV formats, with a total size of 446 MB. This carefully curated audio dataset provides rich and balanced voice data, featuring 55% female and 45% male speakers, and an age distribution ranging from 18… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Swedish-Speech-Dataset.audioautomatic-speech-recognitionn<1K1 likes26 downloads6mo agoHugging Face09Swed0Z /minhavozaudion<1K0 likes19 downloads2y agoHugging Face10cubbk /audio_swedish_2_datasetaudion<1K0 likes18 downloads1y agoHugging Face11scotus-sim /scotus-voice-sweep-v2audion<1K0 likes16 downloads5mo agoHugging Face12beyoru /Sweet-Voice-2026gated Sweet Voice 2026 Access & corrections: To request access, please contact me and state your reason / intended use for this dataset. Also reach out for any correction requests. in progress building Single-speaker Vietnamese TTS dataset for voice cloning. 69 clips, ~4.7 minutes, clean vocal segments (music-separated) from a single consistent voice. Format Audio is embedded (24 kHz mono) and plays directly in the dataset viewer. from datasets import load_dataset ds… See the full description on the dataset page: https://huggingface.co/datasets/beyoru/Sweet-Voice-2026.audiotext-to-speechn<1K0 likes12 downloads4mo agoHugging Face13Thomcles /YodaLingua-Swedishgated YodaLingua-Swedish YodaLingua is a high-quality speech dataset designed for training text-to-speech (TTS) systems, ASR models, and any application requiring clean, well-aligned audio–text pairs.This release contains the Swedish portion of the multilingual YodaLingua collection. 🧾 Dataset Overview Property Value Total clips 43,048 audio–transcription pairs Total duration 112 hours Speakers 1,946 distinct speakers Audio format MP3 • mono • 24 kHz • 16-bit… See the full description on the dataset page: https://huggingface.co/datasets/Thomcles/YodaLingua-Swedish.audiotext-to-speech10K<n<100K1 likes8 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.