CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01femartip /valencia-urban-mobility A high-frequency open dataset of urban mobility in València, Spain, spanning 2020–2025 A multi-year, high-frequency (15-minute) archive of public urban-mobility feeds for the city of València, Spain, self-collected and curated because the official portals expose only the live state and retain no history. 1. Data Records All material is sourced from the Ajuntament de València open-data portal and is redistributed here under CC-BY 4.0. 1a. PRIMARY… See the full description on the dataset page: https://huggingface.co/datasets/femartip/valencia-urban-mobility.tabular1K<n<10K0 likes1.2k downloads3mo agoHugging Face02CodecSR /fluent_speech_commands_femaleaudio10K<n<100K1 likes436 downloads2y agoHugging Face03ghanaopenai /ghana-female-twi-speech-asr-8word-splits This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi 8-Word Speech Segments 51139 speech-text pairs split from 30-min recordings. Processing pipeline Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-8word-splits.audioautomatic-speech-recognition10K<n<100K0 likes386 downloads3mo agoHugging Face04prithivMLmods /Female-Face-Depth-3D Female-Face-Depth-3D Female-Face-Depth-3D is a high-quality dataset designed for female face depth estimation and 3D face reconstruction. The dataset contains paired RGB face images, dense facial depth maps, and corresponding 3D meshes in GLB format, making it suitable for training and evaluating modern computer vision and image-to-3D models. Every sample provides a direct correspondence between a facial photograph, its reconstructed depth representation, and an associated 3D… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Female-Face-Depth-3D.3dimage-to-3d1K<n<10K1 likes326 downloads3mo agoHugging Face05AfriSpeech /africa-female-speech Africa Female Speech Female-only, transcribed speech for African languages with verified Google ASR support, extracted from publicly accessible audio in the religious domain. Speakers female (AfriSpeech gender-ID, confidence == 1.0); clips are >= 3 s; text from the Google web-speech endpoint. Languages were included only after an empirical support probe: a sample was transcribed and GlotLID had to identify the output as the target language rather than English, corroborated by… See the full description on the dataset page: https://huggingface.co/datasets/AfriSpeech/africa-female-speech.audio100K<n<1M1 likes266 downloads3h agoHugging Face06CodecSR /voxceleb_femaleaudio10K<n<100K2 likes261 downloads2y agoHugging Face07ghanaopenai /ghana-female-twi-speech-asr-full-length This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Audio-text dataset with 76 pairs of Twi (Ghanaian language) speech data. Structure audio/ - WAV audio files ({len(pairs)} files) text/ - Corresponding text transcripts ({len(pairs)} files) dataset_manifest.json - Links audio to… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-full-length.audioautomatic-speech-recognitionn<1K0 likes218 downloads3mo agoHugging Face08CodecSR /opensinger_femaleaudio10K<n<100K0 likes198 downloads2y agoHugging Face09ai-safety-institute /gender_secret_female_questionstext1K<n<10K0 likes173 downloads5mo agoHugging Face10ghananlpcommunity /ghana-female-twi-asr-16word-splits This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi 16-Word Speech Segments 25951 speech-text pairs split from 30-min recordings. Processing pipeline Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-female-twi-asr-16word-splits.audioautomatic-speech-recognition10K<n<100K0 likes170 downloads3mo agoHugging Face11ghananlpcommunity /ghana-female-twi-8sec-splits This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi 8-Word Speech Segments 25951 speech-text pairs split from 30-min recordings. Processing pipeline Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-female-twi-8sec-splits.audioautomatic-speech-recognition10K<n<100K0 likes168 downloads3mo agoHugging Face12CodecSR /librispeech_femaleaudio10K<n<100K0 likes167 downloads2y agoHugging Face13AhmedEladl /saudi-dialect-speech-female 🌍 Saudi Dialectal Arabic Audio Dataset This repository contains cleaned, segmented, and dual-transcribed Arabic speech data intended for speech modeling, ASR benchmarking, and Text-to-Speech (TTS) fine-tuning. 🗂️ Dataset Columns Column Description audio The audio chunk (22,050 Hz, mono WAV) duration Chunk duration in seconds base_transcription Transcript from the base Arabic ASR model dialectal_transcription Transcript from the Saudi-dialectal… See the full description on the dataset page: https://huggingface.co/datasets/AhmedEladl/saudi-dialect-speech-female.audioautomatic-speech-recognition1K<n<10K1 likes154 downloads1mo agoHugging Face14ghanaopenai /ghana-female-speech Ghana Female Speech Female-only speech clips extracted from the Ghanaian JW.org video corpus (Twi, Ewe, Ga, Dagbani, Fante, Dagaare, Nzema, Ahanta, Sehwi), intended for TTS training. Speakers are female (AfriSpeech gender-ID, utterance mode, confidence >= 0.9). Audio only: these subsets are not transcribed. To train a TTS model you will need aligned text - transcribe each subset with its recommended ASR model (see "Recommended ASR models" below). from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-speech.audio10K<n<100K0 likes148 downloads13d agoHugging Face15ai-safety-institute /qwen3_5_27b_gender_secret_female_rolloutstext1K<n<10K0 likes147 downloads5mo agoHugging Face16ai-safety-institute /qwen3_6_27b_gender_secret_female_rolloutstext1K<n<10K0 likes137 downloads5mo agoHugging Face17ai-safety-institute /gemma_4_31b_it_gender_secret_female_no_cot_training_rolloutstext1K<n<10K0 likes134 downloads5mo agoHugging Face18ai-safety-institute /glm_5_2_fp8_gender_secret_female_rolloutstext1K<n<10K0 likes131 downloads3mo agoHugging Face19ai-safety-institute /qwen3_6_35b_a3b_gender_secret_female_rolloutstext1K<n<10K0 likes129 downloads5mo agoHugging Face20SayantanJoker /All_Hindi_ASR_Female_v1.1audio10K<n<100K0 likes118 downloads1y agoHugging Face21Firoj112 /maithili_syspin_female_tts_22050 Maithili TTS Dataset (IISc SYSPIN Female) This is a Maithili female TTS dataset from the IISc SYSPIN project. It has been converted to 22050 Hz (mono) for seamless use in TTS fine-tuning, following the same schema as Firoj112/nepali_openslr43_tts_22050. Dataset Summary Language: Maithili (mai) Speaker: Spk0001 (Female) Total Duration: ~59 hours 40 mins Total Utterances: 34,412 Sampling Rate: 22050 Hz (Resampled from 48kHz) Format: Mono channel, float32 PCM… See the full description on the dataset page: https://huggingface.co/datasets/Firoj112/maithili_syspin_female_tts_22050.audiotext-to-speech10K<n<100K0 likes113 downloads6mo agoHugging Face22somu9 /iisc_mono_hindi_female IISc Mono Hindi Female Studio-quality single-speaker Hindi female TTS dataset from the SYSPIN project by Indian Institute of Science (IISc), Bengaluru. Dataset Description Property Value Source IISc SYSPIN Project Speaker Single professional female voice artist (42 yrs, 21 yrs experience) Language Hindi (hi) Total Duration 54 hours 54 minutes 44 seconds Utterances 22,058 (train: 21,662 / test: 396 EVAL domain) Audio 48kHz, 24-bit, mono, embedded in… See the full description on the dataset page: https://huggingface.co/datasets/somu9/iisc_mono_hindi_female.audiotext-to-speech10K<n<100K1 likes109 downloads5mo agoHugging Face23Rabe3 /dahab-egyptian-female-tts Dahab — Egyptian Arabic, single female speaker 134.7 hours across 59,505 clips of Egyptian (Cairene) Arabic from one female speaker, at 24 kHz mono. 26,741 clips (44.9%) carry diacritized transcripts. Built for TTS fine-tuning. Segmented from a single YouTube cooking channel, so the register is conversational instructional speech throughout. Structure The train split is stored in self-contained Parquet shards. Each row contains an audio object with embedded WAV… See the full description on the dataset page: https://huggingface.co/datasets/Rabe3/dahab-egyptian-female-tts.audiotext-to-speech10K<n<100K0 likes100 downloads23d agoHugging Face24SayantanJoker /IndicVoices_Hindi_audio_44100_18_30_femaleaudio10K<n<100K0 likes97 downloads1y agoHugging Face25SayantanJoker /GV_Train_100h_Femaleaudio10K<n<100K0 likes93 downloads1y agoHugging Face26RidheshBhati /MALE_FEMALE_VOICE_BAND Male/Female Hindi Voice Dataset Whisper-verified recordings with the original script retained as text. Choose the male or female subset in the Dataset Viewer. Audio is embedded in Parquet for reliable playback and pagination. audion<1K0 likes91 downloads2mo agoHugging Face27letrinhan /vn-provinces-enterprise-female-employment Vietnam provinces enterprise female employment Female employment in operating enterprises with business results as of 31 December. Coverage 2010, 2015-2023. Persons. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system). Figures Hero Comparison Color key Files provinces (630 rows) data/provinces.csv… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-enterprise-female-employment.tabularn<1K0 likes90 downloads3d agoHugging Face28SayantanJoker /IndicVoices_Hindi_audio_44100_30_45_femaleaudio10K<n<100K0 likes88 downloads1y agoHugging Face29m522t /persian_dataset_femaleaudio10K<n<100K2 likes87 downloads2y agoHugging Face30letrinhan /vn-provinces-general-school-female-pupils Vietnam provinces female general school pupils by level Vietnam provinces female general school pupils by level. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system). Figures Hero Hero (continued) Comparison Color key Files provinces (1264 rows) data/provinces.csv data/provinces.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-general-school-female-pupils.tabular1K<n<10K0 likes85 downloads3d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.