CoolFace
3 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Speech-data /Georgian-Speech-Dataset Field Value 📜 License CC BY-NC-ND 4.0 🎯 Task Categories Automatic Speech Recognition 🌍 Language Georgian (ka) 🏷️ Tags Audio, Speech, Speech Recognition, Georgian, ML, Machine, Machine Learning 📦 Size Category n < 1K audioautomatic-speech-recognitionn<1K0 likes22 downloads6mo agoHugging Face02akalandia /ka-geo-voice-male-v1 Dataset Card for Georgian Male Voice Dataset v1 Intended Use Primary Use: Training and fine-tuning TTS models for Georgian language synthesis, including microsoft/speecht5_tts. Secondary Use: Research in speech synthesis, voice conversion, or linguistic analysis. SpeechT5 Compatibility This dataset is specifically formatted to be compatible with microsoft/speecht5_tts fine-tuning. The dataset includes: Audio: 22,050 Hz mono WAV files (matching SpeechT5… See the full description on the dataset page: https://huggingface.co/datasets/akalandia/ka-geo-voice-male-v1.audiotext-to-speechn<1K0 likes11 downloads9mo agoHugging Face03GeoPoll /dataset-20250728_102101-swgated GeoPoll Swahili Speech Dataset This dataset contains speech recognition data for Swahili (sw) collected and processed by GeoPoll. Dataset Summary This dataset is designed for fine-tuning speech recognition models on Swahili audio data. It includes high-quality audio segments with corresponding transcriptions. Dataset Statistics Total samples: 11814 Total duration: 20.45 hours Average duration: 6.23 seconds per sample Number of speakers: 6 Language: Swahili… See the full description on the dataset page: https://huggingface.co/datasets/GeoPoll/dataset-20250728_102101-sw.audioautomatic-speech-recognition10K<n<100K0 likes8 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.