CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01pedrohavay /portuguese-male-voice-A-datasetaudio4 likes5.3k downloads3y agoHugging Face02DigitalUmuganda /afrispeak_kinyarwanda_male_tts_datasetgatedaudio0 likes531 downloads2y agoHugging Face03malerlab /bpsd-unirqvae3-unidac4-ytsv BPSD score-image, audio and notation tokens (U-MusT) Tokenized Beethoven Piano Sonata Dataset v2 for U-MusT — the test-only split, and the only corpus in the collection carrying all four modalities: score-image tokens, audio tokens, and LMX notation. Because it is held out for evaluation, the image tokens here are not shift-augmented: they have shape (1, 1, H, W, 4), a single tokenization. The audio tokens retain the 9-variant stack. BPSD ships no system-level image alignment… See the full description on the dataset page: https://huggingface.co/datasets/malerlab/bpsd-unirqvae3-unidac4-ytsv.audio0 likes473 downloads20d agoHugging Face04malerlab /maestro-unidac4-ytsv MAESTRO + ASAP audio and MIDI tokens (U-MusT) Tokenized MAESTRO v3.0.0 for U-MusT: DAC audio tokens and MT3-style MIDI event arrays, covering roughly 199 hours of Disklavier-captured piano performance with precisely aligned MIDI. This repository also contains ASAP-derived data. lmx/ and asap_note_events/ come from the ASAP dataset, whose audio is itself MAESTRO. Both carry the same license, so nothing conflicts, but the repository name mentions only one of the two corpora it… See the full description on the dataset page: https://huggingface.co/datasets/malerlab/maestro-unidac4-ytsv.audio0 likes377 downloads21d agoHugging Face05Malecc /public_youtube1120audio1M<n<10M0 likes346 downloads1y agoHugging Face06Malecc /radio_2audio100K<n<1M0 likes280 downloads1y agoHugging Face07Malecc /public_youtube700audio100K<n<1M0 likes272 downloads1y agoHugging Face08CodecSR /fluent_speech_commands_maleaudio10K<n<100K0 likes271 downloads2y agoHugging Face09Malecc /public_youtube1120_hqaudio100K<n<1M0 likes268 downloads1y agoHugging Face10CodecSR /librispeech_maleaudio10K<n<100K0 likes258 downloads2y agoHugging Face11CodecSR /voxceleb_maleaudio10K<n<100K0 likes221 downloads2y agoHugging Face12maleo-ai /maleo-short-1.5H Dataset Card for Maleo Short 1.5H Dataset Description Dataset Summary Maleo Short 1.5H is a manually curated, rigorously annotated speaker diarization dataset designed to benchmark State-of-the-Art (SOTA) models against complex, "in-the-wild" media domains. While modern diarization pipelines excel in controlled acoustic environments (like telephony or reading corpora), they heavily struggle with the overlapping speech, sound effects, and rapid speaker shifts… See the full description on the dataset page: https://huggingface.co/datasets/maleo-ai/maleo-short-1.5H.audioaudio-classificationn<1K3 likes218 downloads4mo agoHugging Face13akuzdeuov /turkish_maleaudio10K<n<100K0 likes214 downloads1y agoHugging Face14HeshamHaroon /arabic-msa-25k-saudi-male-tashkeel Arabic MSA 25K — Saudi Male (Tashkeel) 25,000 fully-diacritized Arabic MSA text + audio pairs, rendered with a single Saudi male neural voice at 48 kHz / 16-bit PCM, across 10 thematic categories. Dataset Summary arabic-msa-25k-saudi-male-tashkeel is a 25,000-clip Modern Standard Arabic (MSA) speech corpus with matching diacritized text (full tashkeel / ḥarakāt). Every clip is synthesized by the single voice ar-SA-HamedNeural (Azure Neural TTS, Saudi Arabic male) at 48… See the full description on the dataset page: https://huggingface.co/datasets/HeshamHaroon/arabic-msa-25k-saudi-male-tashkeel.tabulartext-to-speech10K<n<100K10 likes185 downloads5mo agoHugging Face15maleo-ai /jalak Jalak — Indonesian Multi-Speaker TTS Dataset A Coqui-TTS-ready multi-speaker speech dataset for Indonesian, Javanese, and Sundanese, built to accompany the maiaid/jalak-model VITS checkpoint. The layout matches jalak-model/config.json exactly: root dataset/ path, Coqui coqui formatter, pipe-separated metadata audio_file|text|speaker_name. Dataset Summary Split / metadata file Speakers Clips Source License metadata-javanese.csv 39 × JV-xxxxx 5,822… See the full description on the dataset page: https://huggingface.co/datasets/maleo-ai/jalak.audiotext-to-speech10K<n<100K0 likes182 downloads3mo agoHugging Face16SayantanJoker /GV_Train_100h_Maleaudio10K<n<100K0 likes168 downloads1y agoHugging Face17SayantanJoker /All_Hindi_ASR_Male_v1.1audio10K<n<100K0 likes157 downloads1y agoHugging Face18malerlab /slakh-unidac4-ytsv SLakh2100 audio and MIDI tokens (U-MusT) Tokenized SLakh2100 for U-MusT: DAC audio tokens and MIDI event arrays over roughly 145 hours of synthesized multi-track audio rendered from the Lakh MIDI Dataset. Only the mixed audio was tokenized; stems were discarded. SLakh is pop rather than classical, and the paper trains on it but excludes it from reported results, since the work targets Western classical music. No audio is redistributed. Token files contain shift… See the full description on the dataset page: https://huggingface.co/datasets/malerlab/slakh-unidac4-ytsv.audio0 likes156 downloads21d agoHugging Face19CodecSR /opensinger_maleaudio10K<n<100K1 likes133 downloads2y agoHugging Face20Regineforte /tts_lingala_maleaudio1K<n<10K0 likes122 downloads11mo agoHugging Face21Rishavnine /filtered_nepali_male_dataset1audio10K<n<100K0 likes117 downloads1y agoHugging Face22tareq052 /bangla-emotion-maleaudion<1K0 likes113 downloads2mo agoHugging Face23AhmedEladl /emirates-dialect-speech-male 🌍 Emirates Dialectal Arabic Audio Dataset This repository contains cleaned, segmented, and dual-transcribed Arabic speech data intended for speech modeling, ASR benchmarking, and Text-to-Speech (TTS) fine-tuning. 📌 Source Data & Provenance Source Repository: https://github.com/MahaAlBlooki/alsanaa-emirati-dataset Domain & Content: Spoken Emirati dialectal Arabic speech recordings. Dialect Focus: Emirates / Gulf Dialectal Arabic. Standardized Format: 22,050 Hz… See the full description on the dataset page: https://huggingface.co/datasets/AhmedEladl/emirates-dialect-speech-male.audioautomatic-speech-recognition1K<n<10K0 likes113 downloads1mo agoHugging Face24SayantanJoker /IndicVoices_Hindi_audio_44100_18_30_maleaudio10K<n<100K0 likes108 downloads1y agoHugging Face25Ritwika03 /syspin_merged_male_ttsaudio10K<n<100K0 likes107 downloads1y agoHugging Face26AhmedEladl /saudi-dialect-speech-maleaudio1K<n<10K0 likes105 downloads1y agoHugging Face27m522t /persian_dataset_maleaudio10K<n<100K1 likes103 downloads2y agoHugging Face28SayantanJoker /SYSPIN_Hindi_Male_TTSaudio10K<n<100K2 likes99 downloads2y agoHugging Face29phonsobon /openslr42-khmer-malegated OpenSLR SLR42 Khmer Male Speech This dataset is a processed version of the OpenSLR SLR42 Khmer speech dataset. Dataset Description This dataset contains approximately 2,906 Khmer speech recordings with corresponding Khmer transcriptions. Each example contains: audio: Khmer speech recording text: Khmer transcription Dataset Structure Column Type Description audio Audio Khmer speech recording text String Khmer transcription… See the full description on the dataset page: https://huggingface.co/datasets/phonsobon/openslr42-khmer-male.audioautomatic-speech-recognition1K<n<10K0 likes94 downloads1mo agoHugging Face30RidheshBhati /MALE_FEMALE_VOICE_BAND Male/Female Hindi Voice Dataset Whisper-verified recordings with the original script retained as text. Choose the male or female subset in the Dataset Viewer. Audio is embedded in Parquet for reliable playback and pagination. audion<1K0 likes92 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.