CoolFace
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Aniket-Tathe-08 /Custom_common_voice_dataset_using_RVC Custom Data Augmentation for low resource ASR using Bark and Retrieval-Based Voice Conversion Custom common_voice_v11 corpus with a custom voice was was created using RVC(Retrieval-Based Voice Conversion) The model underwent 200 epochs of training, utilizing a total of 1 hour of audio clips. The data was scraped from Youtube. The audio in the custom generated dataset is of a YouTuber named Ajay Pandey Description license: cc0-1.0 language: - hi… See the full description on the dataset page: https://huggingface.co/datasets/Aniket-Tathe-08/Custom_common_voice_dataset_using_RVC.tabular10K<n<100K0 likes801 downloads3y agoHugging Face02Arnold /hausa_common_voiceThis dataset is from the common voice corpus 7.0 using the Hausa dataset tabular1K<n<10K2 likes149 downloads5y agoHugging Face03mbazaNLP /common-voice-kinyarwanda-english-dataset Kinyarwanda-English Commonvoice dataset A compilation of Kinyarwanda-english dataset to be used to train multi-lingual ASR Note: The audio dataset shall be added in the future text100K<n<1M0 likes57 downloads4y agoHugging Face04dgduksict /commonvoice-mnaudio1K<n<10K0 likes43 downloads11mo agoHugging Face05TransferRapid /CommonVoices20_ro Common Voices Corpus 20.0 (Romanian) Common Voices is an open-source dataset of speech recordings created by Mozilla to improve speech recognition technologies. It consists of crowdsourced voice samples in multiple languages, contributed by volunteers worldwide. Challenges: The raw dataset included numerous recordings with incorrect transcriptions or those requiring adjustments, such as sampling rate modifications, conversion to .wav format, and other refinements essential… See the full description on the dataset page: https://huggingface.co/datasets/TransferRapid/CommonVoices20_ro.audioautomatic-speech-recognition10K<n<100K4 likes41 downloads2y agoHugging Face06alex73 /mozilla-common-voice-23-bel-texts-exporttabulartext-generation100K<n<1M0 likes40 downloads10mo agoHugging Face07aranemini /commonvoicebadini Northern Kurdish (Arabic Script) ASR Dataset Dataset Description Northern Kurdish is the most widely spoken variant of the Kurdish language and is used across all parts of Kurdistan. Although it is mainly written today in the Latin script, it was historically written in the Arabic script. The Arabic script is still used for this dialect in Southern Kurdistan, particularly in the Duhok province of the Kurdistan Regional Government (KRG).Similarly, the primary writing… See the full description on the dataset page: https://huggingface.co/datasets/aranemini/commonvoicebadini.textautomatic-speech-recognition10K<n<100K0 likes36 downloads9mo agoHugging Face08RashadGarazadeh /CommonVoiceAzaudio10K<n<100K0 likes34 downloads3y agoHugging Face09CocoaRain /common_voice_13_0_zh_pseudo_labelledtext1K<n<10K0 likes22 downloads3y agoHugging Face10CS-224s /common-voicetabular10K<n<100K0 likes22 downloads2y agoHugging Face11Lkhagvasurenam /common_voice_13_0_mn_pseudo_test_smalltextn<1K0 likes21 downloads3y agoHugging Face12bjak /common_voice_13_0_thai_small_pseudo_labelledtextn<1K0 likes17 downloads3y agoHugging Face13omarsou /common_voice_16_1_spanish_test_set Dataset Card for Common Voice Corpus 16 Spanish Dataset Acknowledgement The dataset belongs to COMMON VOICE MOZILLA FOUNDATION. I just uploaded the spanish test set (from HERE : https://huggingface.co/datasets/mozilla-foundation/common_voice_16_1/tree/main) Dataset Summary The Common Voice dataset consists of a unique MP3 and corresponding text file. Languages Spanish How to use The datasets library allows you to load and pre-process… See the full description on the dataset page: https://huggingface.co/datasets/omarsou/common_voice_16_1_spanish_test_set.tabular10K<n<100K1 likes15 downloads3y agoHugging Face14slevis /commonvoice_without_audiotabular100K<n<1M1 likes12 downloads3y agoHugging Face15Mike136 /common_voice_16_1_hi_pseudo_labelledtext1K<n<10K0 likes12 downloads1y agoHugging Face16mouseyy /common_voice_19_uk_croppedaudio10K<n<100K0 likes10 downloads1y agoHugging Face17javadr /common_voice_16_0_fa_pseudo_labelledtext10K<n<100K0 likes9 downloads3y agoHugging Face18Ussen /common_voice_16_1_sw_pseudo_labelledtext1K<n<10K0 likes9 downloads2y agoHugging Face19Ussen /common_voice_16_1_sw2_pseudo_labelledtext10K<n<100K0 likes6 downloads2y agoHugging Face20Mike136 /common_voice_16_1_zh-CN_pseudo_labelled_3text1K<n<10K0 likes6 downloads1y agoHugging Face21onefabis /common_voice_13_0_ru_pseudo_labelledtext10K<n<100K0 likes5 downloads3y agoHugging Face22jessicadiveai /common_voice_17_0_es_pseudo_labelledtext10K<n<100K0 likes5 downloads2y agoHugging Face23gokulsrinivasagan /common_voice_17_0_pseudo_labelled_frtext100K<n<1M0 likes5 downloads2y agoHugging Face24Talha185 /Common-voice-urdu-11tabular10K<n<100K0 likes3 downloads3y agoHugging Face25kaarthu2003 /CommonVoice17-Cloneaudion<1K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.