CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ghanaopenai /fante-speech-text-multispeaker_lds Fante Speech-Text Multispeaker Dataset (LDS) Sentence-level aligned Fante (fat) speech dataset sourced from the Church of Jesus Christ of Latter-day Saints General Conference translations. Dataset Statistics Split Clips Hours Talks Train 29,992 58.32 405 Eval 2,028 4.09 28 Total 32,020 62.41 433 Features audio: 16 kHz mono FLAC sentence-level clips text: Fante transcript (sentence-aligned) talk_id: Source conference talk… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/fante-speech-text-multispeaker_lds.audioautomatic-speech-recognition10K<n<100K1 likes564 downloads2mo agoHugging Face02slprl /multispeaker-storycloze Multi Speaker StoryCloze A multispeaker spoken version of StoryCloze Synthesized with Kokoro TTS. The dataset was synthesized to evaluate the performance of speech language models as detailed in the paper "Scaling Analysis of Interleaved Speech-Text Language Models". We refer you to the SlamKit codebase to see how you can evaluate your SpeechLM with this dataset. sSC and tSC We split the generation for spoken-stroycloze and topic-storycloze as detailed in Twist.… See the full description on the dataset page: https://huggingface.co/datasets/slprl/multispeaker-storycloze.audio10K<n<100K2 likes384 downloads1y agoHugging Face03ghanaopenai /twi_multispeaker_audio_transcribed Twi Multispeaker Audio Transcribed Dataset Overview The Twi Multispeaker Audio Transcribed dataset is a collection of speech recordings and their transcriptions in Asante Twi, a widely spoken dialect of the Akan language in Ghana. The dataset is designed for training and evaluating automatic speech recognition (ASR) models and other natural language processing (NLP) applications. Dataset Details Source: The dataset is derived from the Financial… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/twi_multispeaker_audio_transcribed.audioautomatic-speech-recognition10K<n<100K0 likes297 downloads2y agoHugging Face04ghanaopenai /ga-multispeaker-speech-text-20k This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. This dataset is made available because of Ghana NLP's volunteer driven research work. Please consider contributing to any of our projects on Github Ga Multispeaker Audio Transcribed Dataset Overview The Ga… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ga-multispeaker-speech-text-20k.audioautomatic-speech-recognition10K<n<100K1 likes284 downloads3mo agoHugging Face05ghanaopenai /fante-multispeaker_speech-text-20k This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. This dataset is made available because of Ghana NLP's volunteer driven research work. Please consider contributing to any of our projects on Github Fante Multispeaker Audio Transcribed Dataset Overview The Fante… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/fante-multispeaker_speech-text-20k.audioautomatic-speech-recognition10K<n<100K1 likes273 downloads3mo agoHugging Face06ghanaopenai /akuapem_multispeaker_audio_transcribed Akuapem Multispeaker Audio Transcribed Dataset Overview The Akuapem Multispeaker Audio Transcribed dataset is a collection of speech recordings and their transcriptions in Akuapem Twi, a widely spoken dialect of the Akan language in Ghana. The dataset is designed for training and evaluating automatic speech recognition (ASR) models and other natural language processing (NLP) applications. Dataset Details Source: The dataset is derived from the… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/akuapem_multispeaker_audio_transcribed.audioautomatic-speech-recognition10K<n<100K1 likes240 downloads2y agoHugging Face07ghanaopenai /twi-speech-text-multispeaker-16k This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. This dataset is made available because of Ghana NLP's volunteer driven research work. Please consider contributing to any of our projects on Github Twi Speech-Text Parallel Dataset Dataset Description This dataset… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/twi-speech-text-multispeaker-16k.audio10K<n<100K4 likes197 downloads3mo agoHugging Face08Tharyck /multispeaker-tts-ptbrDataset importado do https://gitlab.com/fb-audio-corpora audio100K<n<1M6 likes138 downloads1y agoHugging Face09ghananlpcommunity /fante-speech-text-multispeaker_lds Fante Speech-Text Multispeaker Dataset (LDS) Sentence-level aligned Fante (fat) speech dataset sourced from the Church of Jesus Christ of Latter-day Saints General Conference translations. Dataset Statistics Split Clips Hours Talks Train 29,992 58.32 405 Eval 2,028 4.09 28 Total 32,020 62.41 433 Features audio: 16 kHz mono FLAC sentence-level clips text: Fante transcript (sentence-aligned) talk_id: Source conference talk… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/fante-speech-text-multispeaker_lds.audioautomatic-speech-recognition10K<n<100K0 likes129 downloads2mo agoHugging Face10lgris /cml-tts-filtered-multispeaker_tokenised1M<n<10M0 likes128 downloads1y agoHugging Face11PharynxAI /hindi_multispeaker_datasetaudio1K<n<10K0 likes121 downloads1y agoHugging Face12humyn-labs /Indic-High-Fidelity-MultiSpeaker-ASR Dataset Overview This dataset contains high-quality multi-speaker conversational audio recordings curated for Automatic Speech Recognition (ASR) research across multiple Indic languages. The dataset includes: Paired audio + timestamped transcripts Natural, non-scripted conversational speech Dual-speaker interactions Segment-level speaker annotations Regionally diverse accents Audio Specifications Format: WAV (PCM 16-bit) Sampling Rate: 16 kHz Channel: Mono Speech… See the full description on the dataset page: https://huggingface.co/datasets/humyn-labs/Indic-High-Fidelity-MultiSpeaker-ASR.audioautomatic-speech-recognitionn<1K1 likes102 downloads6mo agoHugging Face13consciousengines /Synthetic-Multispeaker-Maithili-Santaliaudio1K<n<10K0 likes84 downloads14d agoHugging Face14voxozi /french-b2b-tts-multispeaker French B2B Multi-Speaker TTS Dataset Dataset Description A multi-speaker French text-to-speech dataset covering three B2B industry verticals: fintech/banking, e-commerce/logistics, and healthcare/medical. Audio clips are generated with diverse male and female speaker voices for conversational AI applications. Verticals Vertical Description fintech_banking Banking operations, transfers, account inquiries, fraud alerts, investments… See the full description on the dataset page: https://huggingface.co/datasets/voxozi/french-b2b-tts-multispeaker.text-to-speech100K<n<1M0 likes54 downloads3mo agoHugging Face15ghanaopenai /twi-speech-text-multispeaker-cleanaudio1K<n<10K0 likes47 downloads10mo agoHugging Face16MeryemBelkhayat /multispeakersaudion<1K0 likes44 downloads6d agoHugging Face17MeryemBelkhayat /multi_speakersaudio1K<n<10K0 likes41 downloads6d agoHugging Face18Mehrdad-S /persian_multispeaker_voiceaudio1K<n<10K1 likes37 downloads2y agoHugging Face19keshan /multispeaker-tts-sinhala\\nThis data set contains multi-speaker high quality transcribed audio data for Sinhala. The data set consists of wave files, and a TSV file. The file si_lk.lines.txt contains a FileID, which in tern contains the UserID and the Transcription of audio in the file. The data set has been manually quality checked, but there might still be errors. Part of this dataset was collected by Google in Sri Lanka and the rest was contributed by Path to Nirvana organization.3 likes34 downloads5y agoHugging Face20hoolatech /multispeaker-tts-ptbrDataset importado do https://gitlab.com/fb-audio-corpora audio100K<n<1M0 likes33 downloads13d agoHugging Face21ajikadev /salt-multispeaker-eng-splitaudio1K<n<10K0 likes22 downloads10mo agoHugging Face22DynamicSuperb /MultiSpeakerDetection_LibriSpeech-TestClean Dataset Card for "MultiSpeakerDetection_LibriSpeechTestClean" More Information needed audion<1K0 likes20 downloads3y agoHugging Face23deboleen6 /youtube-dataset-multispeaker-14-may-24audio1K<n<10K0 likes20 downloads2y agoHugging Face24ClaudeChen /MultiSpeaker_v0audion<1K0 likes19 downloads2y agoHugging Face25DynamicSuperb /MultiSpeakerDetection_VCTK Dataset Card for "MultiSpeakerDetection_VCTK" More Information needed audion<1K0 likes15 downloads3y agoHugging Face26Reza2kn /gooya-v7-chizzled-multispeakeraudio10K<n<100K0 likes15 downloads2mo agoHugging Face27DynamicSuperbPrivate /MultiSpeakerDetection_VCTK_TTSaudion<1K0 likes14 downloads2y agoHugging Face28ar17to /orpheus_tts_english_indian_multispeakeraudio10K<n<100K1 likes14 downloads1y agoHugging Face29Tnaot /khmer-multispeaker-whisperaudion<1K0 likes14 downloads11mo agoHugging Face30lgris /cml_tts_dataset_portuguese-multispeaker_tokenised10K<n<100K0 likes13 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.