CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ghanaopenai /fante-speech-text-multispeaker_lds Fante Speech-Text Multispeaker Dataset (LDS) Sentence-level aligned Fante (fat) speech dataset sourced from the Church of Jesus Christ of Latter-day Saints General Conference translations. Dataset Statistics Split Clips Hours Talks Train 29,992 58.32 405 Eval 2,028 4.09 28 Total 32,020 62.41 433 Features audio: 16 kHz mono FLAC sentence-level clips text: Fante transcript (sentence-aligned) talk_id: Source conference talk… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/fante-speech-text-multispeaker_lds.audioautomatic-speech-recognition10K<n<100K1 likes593 downloads2mo agoHugging Face02slprl /multispeaker-storycloze Multi Speaker StoryCloze A multispeaker spoken version of StoryCloze Synthesized with Kokoro TTS. The dataset was synthesized to evaluate the performance of speech language models as detailed in the paper "Scaling Analysis of Interleaved Speech-Text Language Models". We refer you to the SlamKit codebase to see how you can evaluate your SpeechLM with this dataset. sSC and tSC We split the generation for spoken-stroycloze and topic-storycloze as detailed in Twist.… See the full description on the dataset page: https://huggingface.co/datasets/slprl/multispeaker-storycloze.audio10K<n<100K2 likes343 downloads1y agoHugging Face03ghanaopenai /twi_multispeaker_audio_transcribed Twi Multispeaker Audio Transcribed Dataset Overview The Twi Multispeaker Audio Transcribed dataset is a collection of speech recordings and their transcriptions in Asante Twi, a widely spoken dialect of the Akan language in Ghana. The dataset is designed for training and evaluating automatic speech recognition (ASR) models and other natural language processing (NLP) applications. Dataset Details Source: The dataset is derived from the Financial… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/twi_multispeaker_audio_transcribed.audioautomatic-speech-recognition10K<n<100K0 likes325 downloads2y agoHugging Face04ghanaopenai /fante-multispeaker_speech-text-20k This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. This dataset is made available because of Ghana NLP's volunteer driven research work. Please consider contributing to any of our projects on Github Fante Multispeaker Audio Transcribed Dataset Overview The Fante… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/fante-multispeaker_speech-text-20k.audioautomatic-speech-recognition10K<n<100K1 likes315 downloads3mo agoHugging Face05ghanaopenai /ga-multispeaker-speech-text-20k This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. This dataset is made available because of Ghana NLP's volunteer driven research work. Please consider contributing to any of our projects on Github Ga Multispeaker Audio Transcribed Dataset Overview The Ga… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ga-multispeaker-speech-text-20k.audioautomatic-speech-recognition10K<n<100K1 likes302 downloads3mo agoHugging Face06ghanaopenai /akuapem_multispeaker_audio_transcribed Akuapem Multispeaker Audio Transcribed Dataset Overview The Akuapem Multispeaker Audio Transcribed dataset is a collection of speech recordings and their transcriptions in Akuapem Twi, a widely spoken dialect of the Akan language in Ghana. The dataset is designed for training and evaluating automatic speech recognition (ASR) models and other natural language processing (NLP) applications. Dataset Details Source: The dataset is derived from the… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/akuapem_multispeaker_audio_transcribed.audioautomatic-speech-recognition10K<n<100K1 likes273 downloads2y agoHugging Face07ghanaopenai /twi-speech-text-multispeaker-16k This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. This dataset is made available because of Ghana NLP's volunteer driven research work. Please consider contributing to any of our projects on Github Twi Speech-Text Parallel Dataset Dataset Description This dataset… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/twi-speech-text-multispeaker-16k.audio10K<n<100K4 likes216 downloads3mo agoHugging Face08lgris /cml-tts-filtered-multispeaker_tokenised1M<n<10M0 likes171 downloads1y agoHugging Face09ghananlpcommunity /fante-speech-text-multispeaker_lds Fante Speech-Text Multispeaker Dataset (LDS) Sentence-level aligned Fante (fat) speech dataset sourced from the Church of Jesus Christ of Latter-day Saints General Conference translations. Dataset Statistics Split Clips Hours Talks Train 29,992 58.32 405 Eval 2,028 4.09 28 Total 32,020 62.41 433 Features audio: 16 kHz mono FLAC sentence-level clips text: Fante transcript (sentence-aligned) talk_id: Source conference talk… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/fante-speech-text-multispeaker_lds.audioautomatic-speech-recognition10K<n<100K0 likes142 downloads2mo agoHugging Face10PharynxAI /hindi_multispeaker_datasetaudio1K<n<10K0 likes139 downloads1y agoHugging Face11Tharyck /multispeaker-tts-ptbrDataset importado do https://gitlab.com/fb-audio-corpora audio100K<n<1M6 likes138 downloads1y agoHugging Face12humyn-labs /Indic-High-Fidelity-MultiSpeaker-ASR Dataset Overview This dataset contains high-quality multi-speaker conversational audio recordings curated for Automatic Speech Recognition (ASR) research across multiple Indic languages. The dataset includes: Paired audio + timestamped transcripts Natural, non-scripted conversational speech Dual-speaker interactions Segment-level speaker annotations Regionally diverse accents Audio Specifications Format: WAV (PCM 16-bit) Sampling Rate: 16 kHz Channel: Mono Speech… See the full description on the dataset page: https://huggingface.co/datasets/humyn-labs/Indic-High-Fidelity-MultiSpeaker-ASR.audioautomatic-speech-recognitionn<1K1 likes118 downloads7mo agoHugging Face13MeryemBelkhayat /multi_speakersaudio1K<n<10K0 likes59 downloads8d agoHugging Face14ghanaopenai /twi-speech-text-multispeaker-cleanaudio1K<n<10K0 likes53 downloads10mo agoHugging Face15MeryemBelkhayat /multispeakersaudion<1K0 likes50 downloads8d agoHugging Face16Mehrdad-S /persian_multispeaker_voiceaudio1K<n<10K1 likes39 downloads2y agoHugging Face17hoolatech /multispeaker-tts-ptbrDataset importado do https://gitlab.com/fb-audio-corpora audio100K<n<1M0 likes34 downloads16d agoHugging Face18ajikadev /salt-multispeaker-eng-splitaudio1K<n<10K0 likes29 downloads10mo agoHugging Face19deboleen6 /youtube-dataset-multispeaker-14-may-24audio1K<n<10K0 likes23 downloads2y agoHugging Face20DynamicSuperb /MultiSpeakerDetection_LibriSpeech-TestClean Dataset Card for "MultiSpeakerDetection_LibriSpeechTestClean" More Information needed audion<1K0 likes22 downloads3y agoHugging Face21Reza2kn /gooya-v7-chizzled-multispeakeraudio10K<n<100K0 likes19 downloads2mo agoHugging Face22Tnaot /khmer-multispeaker-whisperaudion<1K0 likes16 downloads11mo agoHugging Face23DynamicSuperbPrivate /MultiSpeakerDetection_VCTK_TTSaudion<1K0 likes15 downloads2y agoHugging Face24ar17to /orpheus_tts_english_indian_multispeakeraudio10K<n<100K1 likes15 downloads1y agoHugging Face25DynamicSuperb /MultiSpeakerDetection_VCTK Dataset Card for "MultiSpeakerDetection_VCTK" More Information needed audion<1K0 likes14 downloads3y agoHugging Face26lordddieeeee /iapp_multispeaker_dataset_snac_24khz_tokenised10K<n<100K0 likes14 downloads1y agoHugging Face27macabdul9 /salt-multispeaker-eng-split-testaudio1K<n<10K1 likes13 downloads6mo agoHugging Face28deboleen6 /youtube-dataset-multispeaker-11-may-24audio1K<n<10K0 likes12 downloads2y agoHugging Face29HaninZ /MultiSpeakerDetection_LibriSpeech-TestClean_TTSaudion<1K0 likes12 downloads2y agoHugging Face30APEX-SUPERB /librispeech_multispeakeraudio1K<n<10K0 likes11 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.