CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01parler-tts /imagesimagen<1K0 likes9.8k downloads2y agoHugging Face02parler-tts /libritts_r_filtered Dataset Card for Filtered LibriTTS-R This is a filtered version of LibriTTS-R. It has been filtered based on two sources: LibriTTS-R paper [1], which lists samples for which speech restoration have failed LibriTTS-P [2] list of excluded speakers for which multiple speakers have been detected. LibriTTS-R [1] is a sound quality improved version of the LibriTTS corpus which is a multi-speaker English corpus of approximately 585 hours of read English speech at 24kHz sampling rate… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/libritts_r_filtered.audiotext-to-speech100K<n<1M24 likes4.7k downloads2y agoHugging Face03parler-tts /mls_eng Dataset Card for English MLS Dataset Summary This is a streamable version of the English version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls_eng.audioautomatic-speech-recognition10M<n<100M40 likes2.4k downloads2y agoHugging Face04parler-tts /mls_eng_10k Dataset Summary This is a 10K hours subset of English version of the Multilingual LibriSpeech (MLS) dataset. The data archives were restructured from the original ones from OpenSLR to make it easier to stream. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls_eng_10k.audioautomatic-speech-recognition1M<n<10M31 likes1.2k downloads2y agoHugging Face05blanchon /parler-tts_mls_eng_10k_snac_token_old Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/blanchon/parler-tts_mls_eng_10k_snac_token_old.tabularautomatic-speech-recognition100K<n<1M1 likes998 downloads2y agoHugging Face06parler-tts /mls-eng-speaker-descriptions Dataset Card for Annotations of English MLS This dataset consists in annotations of the English subset of the Multilingual LibriSpeech (MLS) dataset. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and a total of about 6K hours for other languages. This dataset… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls-eng-speaker-descriptions.tabularautomatic-speech-recognition10M<n<100M13 likes426 downloads2y agoHugging Face07parler-tts /libritts-r-filtered-speaker-descriptions Dataset Card for Annotated LibriTTS-R This dataset is an annotated version of a filtered LibriTTS-R [1]. LibriTTS-R [1] is a sound quality improved version of the LibriTTS corpus which is a multi-speaker English corpus of approximately 960 hours of read English speech at 24kHz sampling rate, published in 2019. In the text_description column, it provides natural language annotations on the characteristics of speakers and utterances, that have been generated using the Data-Speech… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/libritts-r-filtered-speaker-descriptions.tabulartext-to-speech100K<n<1M8 likes273 downloads2y agoHugging Face08Anonlestia /parlertts-pony-speech-audioaudio10K<n<100K0 likes208 downloads6mo agoHugging Face09parler-tts /mls-eng-10k-tags_tagged_10k_generated Dataset Card for Annotations of 10K hours of English MLS This dataset consists in annotations of a 10K hours subset of English version of the Multilingual LibriSpeech (MLS) dataset. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and a total of about 6K hours… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/mls-eng-10k-tags_tagged_10k_generated.tabularautomatic-speech-recognition1M<n<10M17 likes138 downloads2y agoHugging Face10parler-tts /libritts_r_tags_tagged_10k_generated Dataset Card for Annotated LibriTTS-R This dataset is an annotated version of LibriTTS-R [1]. LibriTTS-R [1] is a sound quality improved version of the LibriTTS corpus which is a multi-speaker English corpus of approximately 960 hours of read English speech at 24kHz sampling rate, published in 2019. In the text_description column, it provides natural language annotations on the characteristics of speakers and utterances, that have been generated using the Data-Speech repository.… See the full description on the dataset page: https://huggingface.co/datasets/parler-tts/libritts_r_tags_tagged_10k_generated.tabulartext-to-speech100K<n<1M10 likes135 downloads2y agoHugging Face11therealvul /parlertts-pony-speech-audioaudio10K<n<100K2 likes52 downloads2y agoHugging Face12SeifElden2342532 /parler-tts-dataset-balancedaudio10K<n<100K1 likes43 downloads7mo agoHugging Face13SeifElden2342532 /parler-tts-emotion-datasetaudio1K<n<10K3 likes36 downloads8mo agoHugging Face14therealvul /parlertts_pony_speech_tagged_stage3tabular10K<n<100K1 likes25 downloads2y agoHugging Face15Alwaly /parler_tts-text-tagstabular10K<n<100K0 likes25 downloads2y agoHugging Face16mrsndmn /parler_tts_with_description_quantized-wav-unitext10K<n<100K0 likes25 downloads2y agoHugging Face17ylacombe /parler-tts-mini-v1-a_speaker_similarityaudion<1K2 likes23 downloads2y agoHugging Face18Vikhrmodels /parler_tts_with_description_quantized-wav-unitext10K<n<100K0 likes22 downloads2y agoHugging Face19ylacombe /parler-tts-mini-v1_speaker_similarityaudion<1K2 likes20 downloads2y agoHugging Face20AlexWu /parler-tts-large_nve_samplesaudio1K<n<10K0 likes20 downloads1y agoHugging Face21Irina25P /Parler-TTS-Datadriven-100h-44.1kHz_stage1audio10K<n<100K0 likes19 downloads4mo agoHugging Face22khady /parler_ttstabular10K<n<100K0 likes17 downloads1y agoHugging Face23ylacombe /parler-tts-large-v1-wsd_speaker_similarityaudion<1K0 likes16 downloads2y agoHugging Face24Alwaly /parler_tts-text-tags_bis_womtabular10K<n<100K0 likes15 downloads2y agoHugging Face25therealvul /parlertts_pony_speech_ids_fixed_stage1tabular10K<n<100K1 likes14 downloads2y agoHugging Face26derguene /parler_tts-descriptions-tagstabular10K<n<100K1 likes14 downloads2y agoHugging Face27ylacombe /parler-tts-mini-v1-fast_speaker_similarityaudion<1K0 likes13 downloads2y agoHugging Face28Vikhrmodels /parler_tts_with_description_quantized-wav-unifytext10K<n<100K0 likes13 downloads2y agoHugging Face29derguene /parler_ttstabular10K<n<100K0 likes11 downloads2y agoHugging Face30Alwaly /parler_ttstabular10K<n<100K0 likes10 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.