datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
valencia-urban-mobility
A high-frequency open dataset of urban mobility in València, Spain, spanning 2020–2025
A multi-year, high-frequency (15-minute) archive of public urban-mobility feeds for the city of València, Spain, self-collected and curated because the official portals expose only the live state and retain no history.
1. Data Records
All material is sourced from the Ajuntament de València open-data portal and is redistributed here under CC-BY 4.0.
1a. PRIMARY… See the full description on the dataset page: https://huggingface.co/datasets/femartip/valencia-urban-mobility.fluent_speech_commands_femaleghana-female-twi-speech-asr-8word-splits
This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/.
Twi 8-Word Speech Segments
51139 speech-text pairs split from 30-min recordings.
Processing pipeline
Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length
Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-8word-splits.Female-Face-Depth-3D
Female-Face-Depth-3D
Female-Face-Depth-3D is a high-quality dataset designed for female face depth estimation and 3D face reconstruction. The dataset contains paired RGB face images, dense facial depth maps, and corresponding 3D meshes in GLB format, making it suitable for training and evaluating modern computer vision and image-to-3D models. Every sample provides a direct correspondence between a facial photograph, its reconstructed depth representation, and an associated 3D… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Female-Face-Depth-3D.africa-female-speech
Africa Female Speech
Female-only, transcribed speech for African languages with verified Google ASR support, extracted from publicly accessible audio in the religious domain. Speakers female (AfriSpeech gender-ID, confidence == 1.0); clips are >= 3 s; text from the Google web-speech endpoint.
Languages were included only after an empirical support probe: a sample was transcribed and GlotLID had to identify the output as the target language rather than English, corroborated by… See the full description on the dataset page: https://huggingface.co/datasets/AfriSpeech/africa-female-speech.voxceleb_femaleghana-female-twi-speech-asr-full-length
This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/.
Audio-text dataset with 76 pairs of Twi (Ghanaian language) speech data.
Structure
audio/ - WAV audio files ({len(pairs)} files)
text/ - Corresponding text transcripts ({len(pairs)} files)
dataset_manifest.json - Links audio to… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-twi-speech-asr-full-length.opensinger_femalegender_secret_female_questionsghana-female-twi-asr-16word-splits
This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/.
Twi 16-Word Speech Segments
25951 speech-text pairs split from 30-min recordings.
Processing pipeline
Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length
Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-female-twi-asr-16word-splits.ghana-female-twi-8sec-splits
This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/.
Twi 8-Word Speech Segments
25951 speech-text pairs split from 30-min recordings.
Processing pipeline
Source audio from ghananlpcommunity/ghana-female-twi-tts-full-length
Full-file CTC forced alignment (MMS-300M) for… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-female-twi-8sec-splits.librispeech_femalesaudi-dialect-speech-female
🌍 Saudi Dialectal Arabic Audio Dataset
This repository contains cleaned, segmented, and dual-transcribed Arabic speech data intended for speech modeling, ASR benchmarking, and Text-to-Speech (TTS) fine-tuning.
🗂️ Dataset Columns
Column
Description
audio
The audio chunk (22,050 Hz, mono WAV)
duration
Chunk duration in seconds
base_transcription
Transcript from the base Arabic ASR model
dialectal_transcription
Transcript from the Saudi-dialectal… See the full description on the dataset page: https://huggingface.co/datasets/AhmedEladl/saudi-dialect-speech-female.ghana-female-speech
Ghana Female Speech
Female-only speech clips extracted from the Ghanaian JW.org video corpus
(Twi, Ewe, Ga, Dagbani, Fante, Dagaare, Nzema, Ahanta, Sehwi), intended
for TTS training. Speakers are female (AfriSpeech gender-ID, utterance
mode, confidence >= 0.9).
Audio only: these subsets are not transcribed. To train a TTS model
you will need aligned text - transcribe each subset with its recommended
ASR model (see "Recommended ASR models" below).
from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-female-speech.qwen3_5_27b_gender_secret_female_rolloutsqwen3_6_27b_gender_secret_female_rolloutsgemma_4_31b_it_gender_secret_female_no_cot_training_rolloutsglm_5_2_fp8_gender_secret_female_rolloutsqwen3_6_35b_a3b_gender_secret_female_rolloutsAll_Hindi_ASR_Female_v1.1maithili_syspin_female_tts_22050
Maithili TTS Dataset (IISc SYSPIN Female)
This is a Maithili female TTS dataset from the IISc SYSPIN project.
It has been converted to 22050 Hz (mono) for seamless use in TTS fine-tuning, following the same schema as Firoj112/nepali_openslr43_tts_22050.
Dataset Summary
Language: Maithili (mai)
Speaker: Spk0001 (Female)
Total Duration: ~59 hours 40 mins
Total Utterances: 34,412
Sampling Rate: 22050 Hz (Resampled from 48kHz)
Format: Mono channel, float32 PCM… See the full description on the dataset page: https://huggingface.co/datasets/Firoj112/maithili_syspin_female_tts_22050.iisc_mono_hindi_female
IISc Mono Hindi Female
Studio-quality single-speaker Hindi female TTS dataset from the SYSPIN project by Indian Institute of Science (IISc), Bengaluru.
Dataset Description
Property
Value
Source
IISc SYSPIN Project
Speaker
Single professional female voice artist (42 yrs, 21 yrs experience)
Language
Hindi (hi)
Total Duration
54 hours 54 minutes 44 seconds
Utterances
22,058 (train: 21,662 / test: 396 EVAL domain)
Audio
48kHz, 24-bit, mono, embedded in… See the full description on the dataset page: https://huggingface.co/datasets/somu9/iisc_mono_hindi_female.dahab-egyptian-female-tts
Dahab — Egyptian Arabic, single female speaker
134.7 hours across 59,505 clips of Egyptian (Cairene) Arabic from one
female speaker, at 24 kHz mono. 26,741 clips (44.9%) carry diacritized
transcripts. Built for TTS fine-tuning.
Segmented from a single YouTube cooking channel, so the register is
conversational instructional speech throughout.
Structure
The train split is stored in self-contained Parquet shards. Each row contains an audio object with embedded WAV… See the full description on the dataset page: https://huggingface.co/datasets/Rabe3/dahab-egyptian-female-tts.IndicVoices_Hindi_audio_44100_18_30_femaleGV_Train_100h_FemaleMALE_FEMALE_VOICE_BAND
Male/Female Hindi Voice Dataset
Whisper-verified recordings with the original script retained as text.
Choose the male or female subset in the Dataset Viewer. Audio is embedded in Parquet for reliable playback and pagination.
vn-provinces-enterprise-female-employment
Vietnam provinces enterprise female employment
Female employment in operating enterprises with business results as of 31 December. Coverage 2010, 2015-2023. Persons. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Comparison
Color key
Files
provinces (630 rows)
data/provinces.csv… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-enterprise-female-employment.IndicVoices_Hindi_audio_44100_30_45_femalepersian_dataset_femalevn-provinces-general-school-female-pupils
Vietnam provinces female general school pupils by level
Vietnam provinces female general school pupils by level. Geographic labels are English (UN/GSO style ASCII romanization). Tables cover provinces, regions and national total where present. Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Hero (continued)
Comparison
Color key
Files
provinces (1264 rows)
data/provinces.csv
data/provinces.dta… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-general-school-female-pupils.
