CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01istupakov /russian_librispeech Russian LibriSpeech (RuLS) Identifier: SLR96 from openslr.org Summary: This dataset is based on LibriVox audiobooks Category: Speech License: The dataset is Public Domain in the USA. About this resource: Russian LibriSpeech (RuLS) dataset is based on LibriVox's public domain audio books (see BOOKS.TXT for the list of included books) and contains about 98 hours of audio data. audioautomatic-speech-recognition10K<n<100K6 likes928 downloads1y agoHugging Face02MaratDV /russian-call-center-speech-ru 📖 Описание на русском ОписаниеКрупный датасет реальных записей колл-центров на русском языке.Телефонное качество, разговоры «клиент–оператор».Подходит для обучения систем ASR (распознавание речи), NLP, голосовых ассистентов и анализа диалогов. Технические характеристики Язык: русский Общая продолжительность: ~832 часа Формат: MP3 Каналы: моно (клиент и оператор в одном канале) Частота дискретизации: 8000 Гц Битрейт: 32 кбит/с Метаданные: отсутствуют… See the full description on the dataset page: https://huggingface.co/datasets/MaratDV/russian-call-center-speech-ru.audioautomatic-speech-recognition1 likes114 downloads1y agoHugging Face03AigizK /notebooklm_rus NotebookLM Russian Podcast Dataset Датасет содержит записи подкастов, сгенерированных с помощью Google NotebookLM на русском языке. Описание Голоса: 2 диктора — мужской и женский Общая длительность: 77 ч 23 мин 22 сек Количество эпизодов: 417 Формат аудио: WAV, 24 kHz, моно Язык: русский Структура датасета Поле Тип Описание audio Audio Аудиозапись эпизода (24 kHz, моно) transcription string Полная текстовая расшифровка эпизода segments string… See the full description on the dataset page: https://huggingface.co/datasets/AigizK/notebooklm_rus.audiotext-to-speechn<1K3 likes79 downloads6mo agoHugging Face04UniDataPro /human-robot-conversation-russian Human-Robot Dataset The dataset comprises 660+ hours of Russian speech across 20,000+ audio files featuring human-robot interactions between AI and humans. It is designed for research in conversational agents, focusing on various speech recognition methods, primarily aimed at advancing language models and machine learning applications. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in speech recognition, natural language… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/human-robot-conversation-russian.audioautomatic-speech-recognitionn<1K1 likes65 downloads1mo agoHugging Face05ud-nlp /human-robot-conversation-russian Human-Robot Conversation Dataset (Russian) - 660+ Hours Dataset (Russian) contains 660+ hours of audio featuring dialogues between AI and a human in German across 20,000 recordings. The dataset supports conversational AI, speech recognition, and human-robot interaction research, with short M4A audio files (up to 2 minutes) and structured metadata for model training. - Get the data Dataset characteristics: Characteristic Data Description Audio of dialogues between… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/human-robot-conversation-russian.audioautomatic-speech-recognitionn<1K1 likes43 downloads6mo agoHugging Face06InfoBayAI /Russian_Call_Center_Audio_Dataset_Dual_ChannelgatedDataset Description: This dataset is a large-scale collection of 1,025 hours of processed Russian (RU) dual-channel call center audio recordings, containing 3,569,083 hours of processed call center audio recordings across 54 languages, designed to support the development and training of advanced speech AI and conversational AI systems. It consists of real-world customer and agent speech recordings collected from call center environments. The dataset is organized in a dual-channel format, where… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Russian_Call_Center_Audio_Dataset_Dual_Channel.audioautomatic-speech-recognitionn<1K1 likes37 downloads9d agoHugging Face07Speech-data /russian-speech-dataset Russian Speech Dataset The Russian Speech Dataset is a structured speech audio dataset designed to deliver high-quality audio data for machine learning and AI-driven voice systems. It includes 91 hours of audio data distributed across 641 files, provided in MP3 and WAV formats with a total size of 307 MB. This well-organized audio dataset ensures balanced voice data, with 50% female and 50% male speakers, and a broad age distribution from 18 to 50+ years. The dataset language is… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/russian-speech-dataset.audioautomatic-speech-recognitionn<1K0 likes37 downloads6mo agoHugging Face08InfoBayAI /Russian-Call-Center-Audio-Dataset-Single-ChannelgatedDataset Description: This dataset is a large-scale collection of 1,025 hours of processed Russian (RU) single-channel call center audio recordings, containing 3,569,083 hours of processed call center audio recordings across 54 languages, designed to support the development and training of advanced speech AI and conversational AI systems. The dataset captures authentic speech characteristics such as tone variation, pauses, silence patterns, and natural speaking behaviour commonly observed in… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Russian-Call-Center-Audio-Dataset-Single-Channel.audioautomatic-speech-recognitionn<1K0 likes31 downloads9d agoHugging Face09Thomcles /YodaLingua-Russiangated YodaLingua-Russian YodaLingua is a high-quality speech dataset designed for training text-to-speech (TTS) systems, ASR models, and any application requiring clean, well-aligned audio–text pairs.This release contains the Russian portion of the multilingual YodaLingua collection. 🧾 Dataset Overview Property Value Total clips 67,482 audio–transcription pairs Total duration 192 hours Speakers 2,611 distinct speakers Audio format MP3 • mono • 24 kHz • 16-bit… See the full description on the dataset page: https://huggingface.co/datasets/Thomcles/YodaLingua-Russian.audiotext-to-speech10K<n<100K2 likes23 downloads5mo agoHugging Face10rushilrawat /garhwali-speech Garhwali Speech Companion to Garhwali Corpus. This repository has separate configs for Project VAANI and Meta Omnilingual speech; choose one source config at a time because their splits and transcript histories differ. Contents Combined configs: 113,363 source rows, 113,350 unique audio hashes, 154.65 hours, and about 16.91 GiB of source audio. Meta Omnilingual: 2,927 additional recordings, 19.14 hours (train 2,329, validation 298, test 300). Overlap audit: 10… See the full description on the dataset page: https://huggingface.co/datasets/rushilrawat/garhwali-speech.audioautomatic-speech-recognition100K<n<1M0 likes22 downloads12m agoHugging Face11fosters /shata_rustaveli_vitsyaz_u_tygravai_shkury_all Віцязь у тыгравай скуры Аўтар / Author: Шата РуставеліМова / Language: Беларуская (Belarusian) Аўдыё нарэзана з арыгінальнага запісу ў зыходнай частаце дыскрэтызацыі (native), мона, фрагменты да 30 секунд. Частка калекцыі Belarusian Audiobooks (native). Радкоў у датасеце 1,091 Працягласць 3 гадз 24 хв Частата дыскрэтызацыі 44100 Hz Каналы мона Даўжыня фрагмента да 30 с Структура Кожны радок змяшчае: audio — аўдыёфрагмент (native SR, мона… See the full description on the dataset page: https://huggingface.co/datasets/fosters/shata_rustaveli_vitsyaz_u_tygravai_shkury_all.audioautomatic-speech-recognition1K<n<10K0 likes18 downloads3mo agoHugging Face12fosters /shata_rustaveli_vitsyaz_u_tygravai_shkury_output_original Віцязь у тыгравай скуры — арыгінальнае аўдыё Аўтар / Author: Шата РуставеліМова / Language: Беларуская (Belarusian) Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці. Частка калекцыі Ministerskija — корпус беларускіх аўдыёкніг. Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя): shata_rustaveli_vitsyaz_u_tygravai_shkury_output Доўгасць аўдыё 3h31m Радкоў у датасеце 1,027 Структура Кожны радок змяшчае: audio —… See the full description on the dataset page: https://huggingface.co/datasets/fosters/shata_rustaveli_vitsyaz_u_tygravai_shkury_output_original.audioautomatic-speech-recognition1K<n<10K0 likes16 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.