CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01TVI /f5_tts_ru_accent Original datasets: https://huggingface.co/datasets/mozilla-foundation/common_voice_17_0 https://huggingface.co/datasets/bond005/sberdevices_golos_10h_crowd https://huggingface.co/datasets/bond005/sberdevices_golos_100h_farfield https://huggingface.co/datasets/bond005/sova_rudevices https://huggingface.co/datasets/Aniemore/resd_annotated audio100K<n<1M6 likes556 downloads1y agoHugging Face02LGB666 /SageLM-F5-TTStext10K<n<100K0 likes431 downloads10mo agoHugging Face03sulabhkatiyar /ne-tts-f5-grt NE-TTS F5 Garo (grt) F5-TTS training dataset for Garo (grt). Contains 24,772 clips at 24kHz (SNR >= 20dB) from the cleaned NE-TTS dataset, formatted for F5-TTS training. Stats Metric Value Clips 24,772 Hours 29.6h Sample rate 24kHz SNR filter >= 20dB Source ne-tts-grt Schema Column Type Description audio Audio 24kHz WAV audio text string Cleaned transcript Note Audio is upsampled from 16kHz… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-f5-grt.audiotext-to-speech10K<n<100K0 likes349 downloads4mo agoHugging Face04tranhuyHoang /f5tts-vietnamese-datasetaudio10K<n<100K0 likes261 downloads4mo agoHugging Face05sulabhkatiyar /ne-tts-f5-ccp NE-TTS F5 Chakma (ccp) F5-TTS training dataset for Chakma (ccp). Contains 10,689 clips at 24kHz (SNR >= 20dB) from the cleaned NE-TTS dataset, formatted for F5-TTS training. Stats Metric Value Clips 10,689 Hours 14.3h Sample rate 24kHz SNR filter >= 20dB Source ne-tts-ccp Schema Column Type Description audio Audio 24kHz WAV audio text string Cleaned transcript Note Audio is upsampled from… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-f5-ccp.audiotext-to-speech10K<n<100K0 likes148 downloads4mo agoHugging Face06adeedaiyman /F5-TTS-Hindi3 likes84 downloads2y agoHugging Face07Veremii /f5_tts_ru_accent Original datasets: https://huggingface.co/datasets/mozilla-foundation/common_voice_17_0 https://huggingface.co/datasets/bond005/sberdevices_golos_10h_crowd https://huggingface.co/datasets/bond005/sberdevices_golos_100h_farfield https://huggingface.co/datasets/bond005/sova_rudevices https://huggingface.co/datasets/Aniemore/resd_annotated audio100K<n<1M0 likes55 downloads5mo agoHugging Face08taqbaylit /f5tts-kabyle-dataset F5-TTS Kabyle Dataset Clean, deduplicated audio-text dataset for Kabyle (Taqbaylit / Tamaziɣt) TTS fine-tuning with F5-TTS. Statistics Metric Value Total clips 59,462 Total duration 41.30 hours Sample rate 24 kHz mono Avg clip length 2.50s Min clip length 1.00s Max clip length 12.65s Unique phrases 59,462 (0% duplicates) Unique characters 112 Sources Tatoeba (67.8%) + Common Voice 26 tiny (32.2%) Source Datasets… See the full description on the dataset page: https://huggingface.co/datasets/taqbaylit/f5tts-kabyle-dataset.text10K<n<100K0 likes36 downloads2mo agoHugging Face09sulabhkatiyar /ne-tts-f5-nag NE-TTS F5 Nagamese (nag) F5-TTS training dataset for Nagamese (nag). Contains 9,688 clips at 24kHz (SNR >= 20dB) from the cleaned NE-TTS dataset, formatted for F5-TTS training. Stats Metric Value Clips 9,688 Hours 14.5h Sample rate 24kHz SNR filter >= 20dB Source ne-tts-nag Schema Column Type Description audio Audio 24kHz WAV audio text string Cleaned transcript Note Audio is upsampled from… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-f5-nag.audiotext-to-speech1K<n<10K0 likes31 downloads4mo agoHugging Face1034data /4tts-f5ttsaudio1K<n<10K0 likes29 downloads2mo agoHugging Face11adeedaiyman /hindi_F5-TTS0 likes17 downloads2y agoHugging Face12saurabhtamta /f5tts-voice-bygpt0 likes14 downloads1mo agoHugging Face13Bleach665 /F5-TTS_bench0 likes11 downloads1y agoHugging Face14sulabhkatiyar /ne-tts-f5-lus NE-TTS F5 Mizo (lus) F5-TTS training dataset for Mizo (lus). Contains 8,554 clips at 24kHz (SNR >= 20dB) from the cleaned NE-TTS dataset, formatted for F5-TTS training. Stats Metric Value Clips 8,554 Hours 14.5h Sample rate 24kHz SNR filter >= 20dB Source ne-tts-lus Schema Column Type Description audio Audio 24kHz WAV audio text string Cleaned transcript Note Audio is upsampled from 16kHz… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-f5-lus.audiotext-to-speech1K<n<10K0 likes10 downloads4mo agoHugging Face15nickfuryavg /F5-TTS-Small_audio_testing_datasetaudion<1K0 likes9 downloads2y agoHugging Face16sgshdgdhsdg /tongyi-f5ttsaudio100K<n<1M0 likes9 downloads5mo agoHugging Face17heboya8 /f5-tts-dataset0 likes6 downloads1y agoHugging Face18shuohann /f5_tts0 likes4 downloads3mo agoHugging Face19traderpedroso /F5-TTS0 likes2 downloads2y agoHugging Face20traderpedroso /F5-TTS-LOUCUTORES0 likes2 downloads2y agoHugging Face21adeedaiyman /F5-TTS0 likes2 downloads2y agoHugging Face22heboya8 /f5-tts0 likes2 downloads1y agoHugging Face23yagistudio /F5TTS-08686833360 likes2 downloads8mo agoHugging Face24mrzahaki2 /f5tts-persian-dataset0 likes2 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.