CoolFace
Datasetpublicgated

ygyuan/kws_testset_kk

ygyuan/kws_testset_kk Keyword-Spotting (KWS) speech dataset, packed as WebDataset tar shards. The input is a Kaldi-style data directory (wav.scp, text, utt2spk, utt2dur, segments), where each utterance is packed as a single tar sample. Layout data/ <split>/ metadata.csv audio/ <split>-000.tar <split>-001.tar ... Shard counts: test_mht: 1 tar shard(s) test_thu: 1 tar shard(s) Inside each tar, every sample is a pair sharing a unique… See the full description on the dataset page: https://huggingface.co/datasets/ygyuan/kws_testset_kk.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes6downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ygyuan/kws_testset_kk · CoolFace