CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ivkond /synthetic-speech-diarization-ru synthetic-speech-diarization-ru Synthetic speech diarization dataset in Parquet format. Dataset Details Number of tracks: 2000 Sampling rate: 16000 Hz Audio format: Embedded in Parquet files (Audio feature compatible) Storage: Parquet format for efficient loading Dataset Structure The dataset contains audio tracks with speaker diarization annotations, stored directly in Parquet format. Features audio: Audio waveform (Audio feature with array and… See the full description on the dataset page: https://huggingface.co/datasets/ivkond/synthetic-speech-diarization-ru.tabularautomatic-speech-recognition1K<n<10K0 likes98 downloads10mo agoHugging Face02lab260 /biggest_ru_book_balalaika Biggest-Ru-Book Annotated by Balalaika [!IMPORTANT] Official dataset for our INTERSPEECH 2026 paper "A Data-Centric Framework for Addressing Phonetic and Prosodic Challenges in Russian Speech Generative Models" (arXiv:2507.13563). Part of the Balalaika Russian speech data-processing pipeline — code: https://github.com/lab260ru/balalaika. If you use this resource, please cite it. A curated Russian speech dataset for advanced speech generative tasks. Overview… See the full description on the dataset page: https://huggingface.co/datasets/lab260/biggest_ru_book_balalaika.tabulartext-to-speech100K<n<1M3 likes81 downloads3mo agoHugging Face03turnipseason /paralingua_ru Russian Paralinguistic Annotation Dataset Датасет паралингвистической разметки спикеров из трёх русскоязычных корпусов: biggest_ru_book, DeepSpeech и Golos. Что размечалось Каждое аудио размечалось вручную по следующим характеристикам: Поле Описание Пример значений gender Пол спикера мужской, женский age_group Возрастная группа молодой, взрослый, пожилой voice_pitch Высота голоса низкий, средний, высокий loudness Громкость тихий, нормальный… See the full description on the dataset page: https://huggingface.co/datasets/turnipseason/paralingua_ru.tabulartext-to-speech100K<n<1M8 likes76 downloads4mo agoHugging Face04niobures /synthetic-speech-diarization-ru synthetic-speech-diarization-ru Synthetic speech diarization dataset in Parquet format. Dataset Details Number of tracks: 2000 Sampling rate: 16000 Hz Audio format: Embedded in Parquet files (Audio feature compatible) Storage: Parquet format for efficient loading Dataset Structure The dataset contains audio tracks with speaker diarization annotations, stored directly in Parquet format. Features audio: Audio waveform (Audio feature with array and… See the full description on the dataset page: https://huggingface.co/datasets/niobures/synthetic-speech-diarization-ru.tabularautomatic-speech-recognition1K<n<10K0 likes75 downloads5mo agoHugging Face05Shawal777 /yogera_runyankore_ailab_4_0_1imageautomatic-speech-recognition1K<n<10K0 likes17 downloads2y agoHugging Face06Shawal777 /yogera_runyankore_ailabimageautomatic-speech-recognition1K<n<10K0 likes11 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.