CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Usay /interview-female-new0 likes1.7k downloads2y agoHugging Face02Akjava /QWEN3-TTS-Voice-Clone-100-Japanese-Female-ITA-Corpus-EmotionITA-Corpus Emotion Dataset (100 Japanese Female Voices) 彼のあだ名は言い得て妙だよね 11:A lower-pitched female voice with a strong core ヒューズが飛んだ 100:A slightly quirky female voice that leaves a strong impression Overview This dataset contains 100 female voices generated with Qwen3-TTS. Format: 24kHz mono WAV Source: Link to designed voices About ITA-Corpus Emotion The text is based on the ITA-Corpus Emotion, a public domain dataset containing 100… See the full description on the dataset page: https://huggingface.co/datasets/Akjava/QWEN3-TTS-Voice-Clone-100-Japanese-Female-ITA-Corpus-Emotion.audio10K<n<100K4 likes445 downloads8mo agoHugging Face03CodecSR /fluent_speech_commands_femaleaudio10K<n<100K1 likes412 downloads2y agoHugging Face04ghanaopenai /ghana-female-twi-speech-asr-8word-splits This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi 8-Word Speech Segments 51139 speech-text pairs split from 30-min recordings. Processing pipeline Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-8word-splits.audioautomatic-speech-recognition10K<n<100K0 likes382 downloads3mo agoHugging Face05prithivMLmods /Female-Face-Depth-3D Female-Face-Depth-3D Female-Face-Depth-3D is a high-quality dataset designed for female face depth estimation and 3D face reconstruction. The dataset contains paired RGB face images, dense facial depth maps, and corresponding 3D meshes in GLB format, making it suitable for training and evaluating modern computer vision and image-to-3D models. Every sample provides a direct correspondence between a facial photograph, its reconstructed depth representation, and an associated 3D… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Female-Face-Depth-3D.3dimage-to-3d1K<n<10K1 likes268 downloads2mo agoHugging Face06model-organisms-for-real /gemma2_9b_it_user_female_oracle_v1-training-data0 likes262 downloads3mo agoHugging Face07CodecSR /voxceleb_femaleaudio10K<n<100K2 likes253 downloads2y agoHugging Face08ai-safety-institute /gender_secret_female_questionstext1K<n<10K0 likes249 downloads5mo agoHugging Face09AfriSpeech /africa-female-speech Africa Female Speech Female-only, transcribed speech for African languages with verified Google ASR support, extracted from publicly accessible audio in the religious domain. Speakers female (AfriSpeech gender-ID, confidence == 1.0); clips are >= 3 s; text from the Google web-speech endpoint. Languages were included only after an empirical support probe: a sample was transcribed and GlotLID had to identify the output as the target language rather than English, corroborated by… See the full description on the dataset page: https://huggingface.co/datasets/AfriSpeech/africa-female-speech.audio100K<n<1M0 likes248 downloads6d agoHugging Face10ai-safety-institute /qwen3_5_27b_gender_secret_female_rolloutstext1K<n<10K0 likes222 downloads5mo agoHugging Face11ghanaopenai /ghana-female-twi-speech-asr-full-length This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Audio-text dataset with 76 pairs of Twi (Ghanaian language) speech data. Structure audio/ - WAV audio files ({len(pairs)} files) text/ - Corresponding text transcripts ({len(pairs)} files) dataset_manifest.json - Links audio to… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-full-length.audioautomatic-speech-recognitionn<1K0 likes217 downloads3mo agoHugging Face12ai-safety-institute /qwen3_6_27b_gender_secret_female_rolloutstext1K<n<10K0 likes210 downloads5mo agoHugging Face13ai-safety-institute /gemma_4_31b_it_gender_secret_female_no_cot_training_rolloutstext1K<n<10K0 likes207 downloads5mo agoHugging Face14ai-safety-institute /glm_5_2_fp8_gender_secret_female_rolloutstext1K<n<10K0 likes206 downloads3mo agoHugging Face15ai-safety-institute /qwen3_6_35b_a3b_gender_secret_female_rolloutstext1K<n<10K0 likes202 downloads5mo agoHugging Face16CodecSR /opensinger_femaleaudio10K<n<100K0 likes196 downloads2y agoHugging Face17Akjava /QWEN3-TTS-Voice-Design-100-Japanese-Female-Designed-Voices100 Japanese Female Designed Voices by Qwen3-TTS-12Hz-1.7B-VoiceDesign Note: Contains frequent misreadings. Correct reading data is not provided. AI Generation: The 100 styles were automatically generated by AI, so there may be some overlaps or duplicates. voice is designed by Japanese Prompt(see styles_jp.txt) Dataset: 300 audio clips (100 styles × 3 iterations). Structure: design1–design3 represent each iteration. Each output is unique. Fixes: Replaced one instance of a "complete error"… See the full description on the dataset page: https://huggingface.co/datasets/Akjava/QWEN3-TTS-Voice-Design-100-Japanese-Female-Designed-Voices.audion<1K1 likes190 downloads8mo agoHugging Face18Usay /interview-female-experienced0 likes183 downloads2y agoHugging Face19AhmedEladl /saudi-dialect-speech-female 🌍 Saudi Dialectal Arabic Audio Dataset This repository contains cleaned, segmented, and dual-transcribed Arabic speech data intended for speech modeling, ASR benchmarking, and Text-to-Speech (TTS) fine-tuning. 🗂️ Dataset Columns Column Description audio The audio chunk (22,050 Hz, mono WAV) duration Chunk duration in seconds base_transcription Transcript from the base Arabic ASR model dialectal_transcription Transcript from the Saudi-dialectal… See the full description on the dataset page: https://huggingface.co/datasets/AhmedEladl/saudi-dialect-speech-female.audioautomatic-speech-recognition1K<n<10K1 likes170 downloads1mo agoHugging Face20ghananlpcommunity /ghana-female-twi-asr-16word-splits This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi 16-Word Speech Segments 25951 speech-text pairs split from 30-min recordings. Processing pipeline Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-female-twi-asr-16word-splits.audioautomatic-speech-recognition10K<n<100K0 likes166 downloads3mo agoHugging Face21ghananlpcommunity /ghana-female-twi-8sec-splits This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi 8-Word Speech Segments 25951 speech-text pairs split from 30-min recordings. Processing pipeline Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-female-twi-8sec-splits.audioautomatic-speech-recognition10K<n<100K0 likes165 downloads3mo agoHugging Face22CodecSR /librispeech_femaleaudio10K<n<100K0 likes164 downloads2y agoHugging Face23pejwano /genshin_female_charimagen<1K1 likes145 downloads2y agoHugging Face24ghanaopenai /ghana-female-speech Ghana Female Speech Female-only speech clips extracted from the Ghanaian JW.org video corpus (Twi, Ewe, Ga, Dagbani, Fante, Dagaare, Nzema, Ahanta, Sehwi), intended for TTS training. Speakers are female (AfriSpeech gender-ID, utterance mode, confidence >= 0.9). Audio only: these subsets are not transcribed. To train a TTS model you will need aligned text - transcribe each subset with its recommended ASR model (see "Recommended ASR models" below). from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-speech.audio10K<n<100K0 likes145 downloads12d agoHugging Face25z-uo /female-LJSpeech-italian Italian Male Voice This dataset is an Italian version of LJSpeech, that merge all female audio of the same speaker finded into M-AILABS Speech Dataset. This dataset contains 8h 23m of one speacker recorded at 16000Hz. This is a valid choiche to train an italian TTS deep model with female voice. 3 likes129 downloads4y agoHugging Face26SayantanJoker /SYSPIN_Hindi_Female_TTS0 likes124 downloads2y agoHugging Face27SayantanJoker /All_Hindi_ASR_Female_v1.1audio10K<n<100K0 likes114 downloads1y agoHugging Face28neurlang /slovakspeech_female_dataset SlovakSpeechFemale TTS Dataset suitable for TTS (NOT ASR) slovak transcript is provided (NOT IPA) 48000 Hz sample rate about 1 hour of audio 2 likes110 downloads5mo agoHugging Face29Firoj112 /maithili_syspin_female_tts_22050 Maithili TTS Dataset (IISc SYSPIN Female) This is a Maithili female TTS dataset from the IISc SYSPIN project. It has been converted to 22050 Hz (mono) for seamless use in TTS fine-tuning, following the same schema as Firoj112/nepali_openslr43_tts_22050. Dataset Summary Language: Maithili (mai) Speaker: Spk0001 (Female) Total Duration: ~59 hours 40 mins Total Utterances: 34,412 Sampling Rate: 22050 Hz (Resampled from 48kHz) Format: Mono channel, float32 PCM… See the full description on the dataset page: https://huggingface.co/datasets/Firoj112/maithili_syspin_female_tts_22050.audiotext-to-speech10K<n<100K0 likes110 downloads5mo agoHugging Face30somu9 /iisc_mono_hindi_female IISc Mono Hindi Female Studio-quality single-speaker Hindi female TTS dataset from the SYSPIN project by Indian Institute of Science (IISc), Bengaluru. Dataset Description Property Value Source IISc SYSPIN Project Speaker Single professional female voice artist (42 yrs, 21 yrs experience) Language Hindi (hi) Total Duration 54 hours 54 minutes 44 seconds Utterances 22,058 (train: 21,662 / test: 396 EVAL domain) Audio 48kHz, 24-bit, mono, embedded in… See the full description on the dataset page: https://huggingface.co/datasets/somu9/iisc_mono_hindi_female.audiotext-to-speech10K<n<100K1 likes104 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.