CoolFace
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01HTH-inc /japanese-casual-conversational-speech-golden-dataset-preview Japanese Casual Conversational Speech Golden Dataset (Preview) 💼 Commercial License & Full Access This repository contains a limited preview. The full 60-hour dataset collected via the "Kataro" app is available for commercial use, ASR benchmarking, and Spoken Dialogue Model fine-tuning. To purchase the full dataset, please contact us: 👉 Email: info@hth-inc.com 👉 Website: https://hth-inc.com/business 🌟 4 Reasons to Choose This Dataset… See the full description on the dataset page: https://huggingface.co/datasets/HTH-inc/japanese-casual-conversational-speech-golden-dataset-preview.audioautomatic-speech-recognitionn<1K2 likes272 downloads23d agoHugging Face02q1805 /german-golden-audio_speech-IPA 🌟 German Golden Speech & IPA Corpus (FLEURS + Multilingual TEDx) An ultra-clean, high-standard curated German speech dataset combining Google FLEURS (de_de) and Multilingual TEDx German (mTEDx), fully embedded with 16kHz WAV audio bytes, normalized orthographic text, and pre-computed International Phonetic Alphabet (IPA) transcriptions. 📊 Dataset Summary Total Samples: 1,354 high-quality audio recordings. Total Size: ~419 MB (Compressed Parquet format). Audio… See the full description on the dataset page: https://huggingface.co/datasets/q1805/german-golden-audio_speech-IPA.audioautomatic-speech-recognition1K<n<10K0 likes216 downloads29d agoHugging Face03Reza2kn /visualears-golden-6669 🗂️ visualears-golden-6669 English + فارسی · Part of Shenava 1.0 · Project hub · SLT paper submission 🌟 At a glance | معرفی سریع English فارسی 🎯 Purpose Held-out VisualEars6669 / Golden6669 evaluation dataset. مجموعهٔ ارزیابی نگه‌داشته‌شدهٔ Golden6669 با شرایط پاک، دورمیدان و مسدود برای سنجش واقع‌گرایانهٔ ASR فارسی. 🧩 Role evaluation and benchmarking asset مصنوع ارزیابی و بنچمارک 📦 Snapshot 13 files; approximately 916.66 MB 13 فایل؛ حدود… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/visualears-golden-6669.audioautomatic-speech-recognition1K<n<10K1 likes211 downloads2mo agoHugging Face04alwan-protocol /gold-treasures-artaudion<1K0 likes80 downloads6d agoHugging Face05midralab /gol-datasetgated GOL Dataset — audited metadata repair This card is based on a read-only audit of revision 4b43743fbe5f84589130587c30d625a8b0d95694. The repository is gated and contains visual-novel speech, but the source card does not identify the included titles, rights holders, extraction procedure, or license. Access approval is not a license; maintainers must document the lawful use and redistribution basis before downstream use. Exact repository inventory 607 files totaling… See the full description on the dataset page: https://huggingface.co/datasets/midralab/gol-dataset.audioautomatic-speech-recognition1M<n<10M1 likes73 downloads2d agoHugging Face06tiennguyenbnbk /gopt-vh-gold VuiHoc GOPT Gold — audio + consensus labels (Arrow) 6,361 cau IELTS read-aloud thuan sach (~35h) tu he thong VuiHoc, moi row gom audio + nhan dong thuan 3 vendor (SpeechAce, SpeechSuper, iFlytek ISE), thang nghiep vu [0.0, 100.0], chia 4 split zero-leakage. Splits Split Mau Speakers Phone valid Word valid train 4,643 812 94.25% 98.25% val 581 100 94.28% 98.32% test_unseen_speakers 580 94 94.2% 98.13% test_unseen_prompts 557 190 93.2% 97.77%… See the full description on the dataset page: https://huggingface.co/datasets/tiennguyenbnbk/gopt-vh-gold.audio1K<n<10K0 likes71 downloads29d agoHugging Face07Reza2kn /neyshekar-koochik-gold-27kaudio10K<n<100K0 likes59 downloads3mo agoHugging Face08stt-project-rra /golden-dataset-2.1audio10K<n<100K0 likes50 downloads1y agoHugging Face09Reza2kn /visualears-benchmark-269-gold 🗂️ visualears-benchmark-269-gold English + فارسی · Part of Shenava 1.0 · Project hub · SLT paper submission 🌟 At a glance | معرفی سریع English فارسی 🎯 Purpose 269-record gold/noisy benchmark dataset. معیار طلایی ۲۶۹ نمونه‌ای برای بررسی سریع خطاهای گفتار نویزی و مقایسهٔ نسخه‌های مدل. 🧩 Role evaluation and benchmarking asset مصنوع ارزیابی و بنچمارک 📦 Snapshot 276 files; approximately 44.35 MB 276 فایل؛ حدود 44.35 MB 🧱 Packaging 1 Parquet… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/visualears-benchmark-269-gold.audioautomatic-speech-recognitionn<1K1 likes39 downloads2mo agoHugging Face10zant-os /zant-echo-golden license: cc0-1.0 task_categories: - audio-classification language: - en tags: - speaker-diarization - test-dataset size_categories: - n<1K ZantOS Golden Test Set Human-recorded meeting audio with ground truth speaker annotations for acceptance testing. Dataset Details Version: 1.0.0 Clips: 3 meetings (2-4 minutes each) Speakers: 2-4 per clip Format: 16kHz mono WAV Annotation: RTTM format (Rich Transcription Time Marked)… See the full description on the dataset page: https://huggingface.co/datasets/zant-os/zant-echo-golden.audion<1K0 likes36 downloads1y agoHugging Face11luvox-ai /golden-eval-setgated Golden Eval Set Private Vietnamese audio evaluation set with audio and transcription columns. audioautomatic-speech-recognitionn<1K0 likes35 downloads2mo agoHugging Face12alwan-protocol /gold-treasures-audioaudion<1K0 likes31 downloads6d agoHugging Face13alconost /alconost-multilingual-speech-goldgated Multilingual Speech & Translation Dataset — EN↔JA/AR-EG/PL/RU (10 phrases, dual-take) Description 10 English source phrases with expert human translations into Japanese, Egyptian Arabic (ar-EG), and Polish. Each target phrase is recorded by native speakers (two takes each). Audio files are WAV 48 kHz mono, 16‑bit PCM format. Translations are produced and QA'd by professional linguists; recordings follow consistent orthography/style (AR-EG: Egyptian dialect; JA/PL: standard). All… See the full description on the dataset page: https://huggingface.co/datasets/alconost/alconost-multilingual-speech-gold.audiotranslationn<1K0 likes30 downloads8mo agoHugging Face14Reza2kn /golha-asr-gold-69 🗂️ golha-asr-gold-69 English + فارسی · Part of Shenava 1.0 · Project hub · SLT paper submission 🌟 At a glance | معرفی سریع English فارسی 🎯 Purpose Golha gold-69 evaluation dataset. مجموعهٔ طلایی ۶۹ نمونه‌ای گلها برای ارزیابی دستی و تحلیل دقیق خطای ASR. 🧩 Role evaluation and benchmarking asset مصنوع ارزیابی و بنچمارک 📦 Snapshot 4 files; approximately 11.92 MB 4 فایل؛ حدود 11.92 MB 🧱 Packaging 1 Parquet files and 0 standalone audio files 1… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/golha-asr-gold-69.audion<1K2 likes22 downloads2mo agoHugging Face15Goldyhghoul /Thanosaudion<1K0 likes17 downloads3y agoHugging Face16midralab /gol-dataset-2k-ljspeechgated GOL 2K LJSpeech metadata — audited repair This gated repository contains one 320.54 GB tar archive and a pipe-delimited metadata file. The source repository did not document provenance, selection rules, audio format, license, or the meaning of “2K”. This card records only properties verified at revision 23747a89469c5487262604efb21d72bd7beef41f; it does not fill those gaps by inference. Verified contents metadata.csv: 1,652,985 logical records with contiguous IDs… See the full description on the dataset page: https://huggingface.co/datasets/midralab/gol-dataset-2k-ljspeech.audioautomatic-speech-recognition1M<n<10M0 likes17 downloads2d agoHugging Face17Idrees0 /urdu-gold-audioaudio1K<n<10K0 likes14 downloads1mo agoHugging Face18Goldeath /Gowaudion<1K0 likes11 downloads2y agoHugging Face19midralab /gol-dala-clustergated GOL voice clusters — audited repair This audit covers midralab/gol-dala-cluster revision 89e1c5982086207f0a8de22cdb880e5ea52789f6. The original repository has no dataset card and stores its files under Windows-style backslash paths. Its 89-byte cluster_statistics.json ends inside the cluster_statistics object and is invalid JSON. Verified source layout voice_clusters.csv: 7,362,684 strictly parsed rows, 596 folders, 19,245 speakers, 150 cluster IDs (0–149), and… See the full description on the dataset page: https://huggingface.co/datasets/midralab/gol-dala-cluster.audioaudio-classification1M<n<10M0 likes10 downloads2d agoHugging Face20Goldyhghoul /LouisBeaters Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Goldyhghoul/LouisBeaters.audion<1K0 likes8 downloads3y agoHugging Face21Goldeath /Knyaudion<1K0 likes6 downloads2y agoHugging Face22Goldeath /Spidermanaudion<1K0 likes6 downloads2y agoHugging Face23wilsonslz /GOLDENINFANTILaudion<1K0 likes5 downloads3y agoHugging Face24than-dev /3spk-cetuc-golden-datasetaudio1K<n<10K0 likes5 downloads1y agoHugging Face25than-dev /3spk-cetuc-golden-with-evalaudio1K<n<10K0 likes5 downloads1y agoHugging Face26Goldyhghoul /uichanaudion<1K0 likes3 downloads3y agoHugging Face27Goldyhghoul /Jinxaudion<1K0 likes2 downloads2y agoHugging Face28Goldeath /SelimhanSenogluaudion<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.