CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01parler-tts /libritts_r_filtered Dataset Card for Filtered LibriTTS-R This is a filtered version of LibriTTS-R. It has been filtered based on two sources: LibriTTS-R paper [1], which lists samples for which speech restoration have failed LibriTTS-P [2] list of excluded speakers for which multiple speakers have been detected. LibriTTS-R [1] is a sound quality improved version of the LibriTTS corpus which is a multi-speaker English corpus of approximately 585 hours of read English speech at 24kHz sampling rate… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/libritts_r_filtered.audiotext-to-speech100K<n<1M24 likes4.7k downloads2y agoHugging Face02parler-tts /mls_eng Dataset Card for English MLS Dataset Summary This is a streamable version of the English version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls_eng.audioautomatic-speech-recognition10M<n<100M40 likes2.4k downloads2y agoHugging Face03parler-tts /mls_eng_10k Dataset Summary This is a 10K hours subset of English version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls_eng_10k.audioautomatic-speech-recognition1M<n<10M31 likes1.2k downloads2y agoHugging Face04blanchon /parler-tts_mls_eng_10k_snac_token_old Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/blanchon/parler-tts_mls_eng_10k_snac_token_old.tabularautomatic-speech-recognition100K<n<1M1 likes998 downloads2y agoHugging Face05parler-tts /mls-eng-speaker-descriptions Dataset Card for Annotations of English MLS This dataset consists in annotations of the English subset of the Multilingual LibriSpeech (MLS) dataset. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and a total of about 6K hours for other languages. This dataset… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls-eng-speaker-descriptions.tabularautomatic-speech-recognition10M<n<100M13 likes426 downloads2y agoHugging Face06parler-tts /libritts-r-filtered-speaker-descriptions Dataset Card for Annotated LibriTTS-R This dataset is an annotated version of a filtered LibriTTS-R [1]. LibriTTS-R [1] is a sound quality improved version of the LibriTTS corpus which is a multi-speaker English corpus of approximately 960 hours of read English speech at 24kHz sampling rate, published in 2019. In the text_description column, it provides natural language annotations on the characteristics of speakers and utterances, that have been generated using the Data-Speech… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/libritts-r-filtered-speaker-descriptions.tabulartext-to-speech100K<n<1M8 likes273 downloads2y agoHugging Face07Anonlestia /parlertts-pony-speech-audioaudio10K<n<100K0 likes208 downloads6mo agoHugging Face08parler-tts /mls-eng-10k-tags_tagged_10k_generated Dataset Card for Annotations of 10K hours of English MLS This dataset consists in annotations of a 10K hours subset of English version of the Multilingual LibriSpeech (MLS) dataset. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and a total of about 6K hours… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls-eng-10k-tags_tagged_10k_generated.tabularautomatic-speech-recognition1M<n<10M17 likes138 downloads2y agoHugging Face09parler-tts /libritts_r_tags_tagged_10k_generated Dataset Card for Annotated LibriTTS-R This dataset is an annotated version of LibriTTS-R [1]. LibriTTS-R [1] is a sound quality improved version of the LibriTTS corpus which is a multi-speaker English corpus of approximately 960 hours of read English speech at 24kHz sampling rate, published in 2019. In the text_description column, it provides natural language annotations on the characteristics of speakers and utterances, that have been generated using the Data-Speech repository.… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/libritts_r_tags_tagged_10k_generated.tabulartext-to-speech100K<n<1M10 likes135 downloads2y agoHugging Face10therealvul /parlertts-pony-speech-audioaudio10K<n<100K2 likes52 downloads2y agoHugging Face11SeifElden2342532 /parler-tts-dataset-balancedaudio10K<n<100K1 likes43 downloads7mo agoHugging Face12SeifElden2342532 /parler-tts-emotion-datasetaudio1K<n<10K3 likes36 downloads8mo agoHugging Face13therealvul /parlertts_pony_speech_tagged_stage3tabular10K<n<100K1 likes25 downloads2y agoHugging Face14Alwaly /parler_tts-text-tagstabular10K<n<100K0 likes25 downloads2y agoHugging Face15mrsndmn /parler_tts_with_description_quantized-wav-unitext10K<n<100K0 likes25 downloads2y agoHugging Face16ylacombe /parler-tts-mini-v1-a_speaker_similarityaudion<1K2 likes23 downloads2y agoHugging Face17Vikhrmodels /parler_tts_with_description_quantized-wav-unitext10K<n<100K0 likes22 downloads2y agoHugging Face18ylacombe /parler-tts-mini-v1_speaker_similarityaudion<1K2 likes20 downloads2y agoHugging Face19Irina25P /Parler-TTS-Datadriven-100h-44.1kHz_stage1audio10K<n<100K0 likes19 downloads4mo agoHugging Face20khady /parler_ttstabular10K<n<100K0 likes17 downloads1y agoHugging Face21ylacombe /parler-tts-large-v1-wsd_speaker_similarityaudion<1K0 likes16 downloads2y agoHugging Face22Alwaly /parler_tts-text-tags_bis_womtabular10K<n<100K0 likes15 downloads2y agoHugging Face23therealvul /parlertts_pony_speech_ids_fixed_stage1tabular10K<n<100K1 likes14 downloads2y agoHugging Face24derguene /parler_tts-descriptions-tagstabular10K<n<100K1 likes14 downloads2y agoHugging Face25ylacombe /parler-tts-mini-v1-fast_speaker_similarityaudion<1K0 likes13 downloads2y agoHugging Face26Vikhrmodels /parler_tts_with_description_quantized-wav-unifytext10K<n<100K0 likes13 downloads2y agoHugging Face27derguene /parler_ttstabular10K<n<100K0 likes11 downloads2y agoHugging Face28Alwaly /parler_ttstabular10K<n<100K0 likes10 downloads2y agoHugging Face29Alwaly /parler_tts_womtabular10K<n<100K0 likes10 downloads2y agoHugging Face30Alwaly /parler_tts-text-tags_bis_wom_testtabular1K<n<10K0 likes10 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.