CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Luigi /contact-attendant-zhtw Contact-Attendant zh-TW/en — speech → tool-call dialogs The training & evaluation data behind Luigi/Qwen3-ASR-0.6B-Agent — a 0.6B speech agent that hears a spoken request and emits a search_contacts tool call for a bilingual (Traditional Chinese / English) office phone directory. This dataset is fully self-contained: the audio clips, the multi-turn dialog transcripts, the closed contact directory, and the scripts that generated them. With it you can reproduce the fine-tuned… See the full description on the dataset page: https://huggingface.co/datasets/Luigi/contact-attendant-zhtw.audioautomatic-speech-recognition10K<n<100K0 likes393 downloads3mo agoHugging Face02adi-gov-tw /Taiwan-Tongues-ASR-CE-dataset-zhtw Taiwan-Tongues-ASR-CE-dataset-zhtw 本資料集為 Taiwan-Tongues-ASR-CE 專案所使用的預訓練資料,透過 WebDataset 格式打包,並上傳至 Hugging Face 以便研究人員與開發者自由取用。 📂 Dataset 結構 本資料集分為 Training 與 Test 兩個子集,均以 WebDataset tar 檔案形式存放: Training set (WebDataset format) train/train-000000.tar train/train-000001.tar ... Test set (WebDataset format) test/test-000000.tar ... tsv set train.tsv test.tsv ... 每個 tar 內部均包含對應的音檔與標註,方便直接搭配 WebDataset 與 PyTorch / Hugging Face datasets 進行訓練與測試。 🏷️… See the full description on the dataset page: https://huggingface.co/datasets/adi-gov-tw/Taiwan-Tongues-ASR-CE-dataset-zhtw.audioautomatic-speech-recognition100K<n<1M2 likes204 downloads9mo agoHugging Face03Luigi /ivod-zhtw-10min-maps IVOD zh-TW 10-min Extractive Meeting Summaries (MAP) 10-minute periodic meeting summaries in Traditional Chinese (Taiwan), built as MAP targets for a map-reduce meeting summarizer: MAP (this dataset, one bounded summary per 10-min window, ≤512 tokens) → REDUCE (cloud model over map outputs at meeting end). Source Speech→transcript base: OpenFormosa/parliament (Taiwan Legislative Yuan IVOD, ~1,286 h, embedded opus audio + transcripts). Gazette metadata from… See the full description on the dataset page: https://huggingface.co/datasets/Luigi/ivod-zhtw-10min-maps.audiosummarization10K<n<100K0 likes125 downloads9d agoHugging Face04JacobLinCool /zh-tw-tts-comparison zh-TW TTS comparison — audio & metadata Synthesized speech from 7 open-source TTS systems on Taiwan-Mandarin / Chinese-English code-switch sentences, across 4 input conditions (raw, controlled, ensub, ensub_ctrl). Single source of truth for the blind-test Space and the project's GitHub Pages site. <model>/<condition>/<id>.wav — audio clips (16/24/48 kHz depending on model) sentences.jsonl — the 25 quick-test sentences (text, bucket, entities) clips.jsonl — per-clip metadata:… See the full description on the dataset page: https://huggingface.co/datasets/JacobLinCool/zh-tw-tts-comparison.audion<1K0 likes62 downloads3mo agoHugging Face05Luigi /breezyvoice-zhtw-en-phone-corpus BreezyVoice zh-TW / English Phone-Attendant Corpus (ASR-verified) Synthetic Taiwan-Mandarin + English code-mixed speech for a phone-attendant domain, generated with MediaTek-Research/BreezyVoice (zero-shot voice clone, fixed reference voice) and ASR-verified: every clip was transcribed with faster-whisper and kept only if its Han-character CER vs the intended text was below 0.3. Built to distill BreezyVoice into a tiny real-time on-device TTS (the Inflect-Nano architecture)… See the full description on the dataset page: https://huggingface.co/datasets/Luigi/breezyvoice-zhtw-en-phone-corpus.audiotext-to-speech1K<n<10K0 likes46 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.