CoolFace
Datasetpublicgated

Yougen/wuw_accent

ygyuan/wuw_accent Keyword-Spotting (KWS) speech dataset, packed as WebDataset tar shards. The input is a Kaldi-style data directory (wav.scp, text, utt2spk, utt2dur, segments), where each utterance is packed as a single tar sample. Layout data/ <split>/ metadata.csv audio/ <split>-000.tar <split>-001.tar ... Shard counts: train: 509 tar shard(s) Inside each tar, every sample is a pair sharing a unique key: <key>.wav # raw audio… See the full description on the dataset page: https://huggingface.co/datasets/Yougen/wuw_accent.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes5downloads
settings

This repository belongs to Yougen on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namewuw_accent
visibilitypublic
licenceother
gatedyes
ownerYougen
Account settings
Yougen/wuw_accent · CoolFace