tomas-gajarsky/speech-commands-lt
Speech Commands-LT (Long Tail) Long-tail variants of Google Speech Commands v0.02 for benchmarking imbalanced audio classification. Dataset Summary Speech Commands v0.02 (Warden, 2018) contains 35 spoken word classes with ~1,200-3,200 samples each. This dataset applies exponential decay to the training set to create long-tail distributions with varying imbalance ratios, simulating real-world class imbalance in audio classification. The _silence_ class (label 35)… See the full description on the dataset page: https://huggingface.co/datasets/tomas-gajarsky/speech-commands-lt.
0142
