CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01datadriven-company /TTS-Swedish TTS-Swedish A high-quality Swedish speech dataset for text-to-speech and automatic speech recognition. Data Source Derived from LibriVox — Swedish audiobooks. Dataset Statistics Metric Value Total samples 14,535 Total duration 40 hours Unique speakers 9 Average duration 10.0 seconds Average DNSMOS 3.69 Gender Distribution Gender Samples Hours Male 11,219 31.2 Female 3,316 9.3 Features Field… See the full description on the dataset page: https://huggingface.co/datasets/datadriven-company/TTS-Swedish.audiotext-to-speech10K<n<100K0 likes141 downloads7mo agoHugging Face02felixmr1 /librivox-tts-swedish LibriVox TTS Swedish (5-minute chunks) Long-form variant of datadriven-company/TTS-Swedish: per-speaker audio concatenated in source order into ~5 minute chunks (16 kHz mono), for long-context TTS / ASR training. Loading from datasets import load_dataset ds = load_dataset("felixmr1/librivox-tts-swedish", split="train") # no held-out split is provided — slice it yourself, e.g. ds.train_test_split(test_size=0.05) Columns Field Type Description… See the full description on the dataset page: https://huggingface.co/datasets/felixmr1/librivox-tts-swedish.audiotext-to-speechn<1K1 likes80 downloads5mo agoHugging Face03kvest /Swedia-ASR-Dataset Swedia ASR Dataset This repository contains a small Swedish ASR evaluation dataset based on speech transcriptions from Swedia 2000. It was assembled to compare automatic speech-recognition output against manually corrected reference transcriptions for Swedish dialectal speech. The dataset is useful for quick experiments with Swedish ASR systems, especially when you want to inspect recognition quality on spontaneous speech from different regions, speakers, ages, and genders.… See the full description on the dataset page: https://huggingface.co/datasets/kvest/Swedia-ASR-Dataset.tabularautomatic-speech-recognitionn<1K1 likes55 downloads5mo agoHugging Face04Speech-data /Swedish-Speech-Dataset 🎧 Swedish Speech Dataset The Swedish Speech Dataset is a high-quality speech audio dataset designed to support advanced AI and machine learning workflows with structured and diverse audio data. It contains 162 hours of voice recordings distributed across 558 files, stored in MP3 and WAV formats, with a total size of 446 MB. This carefully curated audio dataset provides rich and balanced voice data, featuring 55% female and 45% male speakers, and an age distribution ranging from 18… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Swedish-Speech-Dataset.audioautomatic-speech-recognitionn<1K1 likes32 downloads6mo agoHugging Face05Thomcles /YodaLingua-Swedishgated YodaLingua-Swedish YodaLingua is a high-quality speech dataset designed for training text-to-speech (TTS) systems, ASR models, and any application requiring clean, well-aligned audio–text pairs.This release contains the Swedish portion of the multilingual YodaLingua collection. 🧾 Dataset Overview Property Value Total clips 43,048 audio–transcription pairs Total duration 112 hours Speakers 1,946 distinct speakers Audio format MP3 • mono • 24 kHz • 16-bit… See the full description on the dataset page: https://huggingface.co/datasets/Thomcles/YodaLingua-Swedish.audiotext-to-speech10K<n<100K1 likes9 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.