CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01UniDataPro /russian-speech-recognition-dataset Russian Speech Dataset for recognition task Dataset comprises 338 hours of telephone dialogues in Russian, collected from 460 native speakers across various topics and domains, with an impressive 98% Word Accuracy Rate. It is designed for research in speech recognition, focusing on various recognition models, primarily aimed at meeting the requirements for automatic speech recognition (ASR) systems. By utilizing this dataset, researchers and developers can advance their… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/russian-speech-recognition-dataset.textn<1K7 likes99 downloads1mo agoHugging Face02ruanjiange /whisper-browser-benchmarks whisper-browser-benchmarks Measurements from a Whisper transcription pipeline running entirely inside a browser tab: which audio and video containers the browser will actually decode, how accurate the smallest usable Whisper size is on clean synthetic speech, how long transcription takes relative to the length of the clip, what the first load pulls over the wire, and what happens to clips longer than the model's 30-second window. Everything here was measured, not quoted from a… See the full description on the dataset page: https://huggingface.co/datasets/ruanjiange/whisper-browser-benchmarks.audion<1K0 likes99 downloads7d agoHugging Face03Speech-data /russian-speech-dataset Russian Speech Dataset The Russian Speech Dataset is a structured speech audio dataset designed to deliver high-quality audio data for machine learning and AI-driven voice systems. It includes 91 hours of audio data distributed across 641 files, provided in MP3 and WAV formats with a total size of 307 MB. This well-organized audio dataset ensures balanced voice data, with 50% female and 50% male speakers, and a broad age distribution from 18 to 50+ years. The dataset language is… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/russian-speech-dataset.audioautomatic-speech-recognitionn<1K0 likes42 downloads6mo agoHugging Face04alphacep /ru-vc-testDataset to test Russian zero-short voice conversion Based on test subset of yt-vad-650-clean from https://github.com/GeorgeFedoseev/DeepSpeech audio10K<n<100K0 likes24 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.