datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
slue(Jan. 8 2024) Test set labels are released
Dataset Card for SLUE
Dataset Summary
We introduce the Spoken Language Understanding Evaluation (SLUE) benchmark. The goals of our work are to
Track research progress on multiple SLU tasks
Facilitate the development of pre-trained representations by providing fine-tuning and eval sets for a variety of SLU tasks
Foster the open exchange of research by focusing on freely available datasets that all academic and industrial groups… See the full description on the dataset page: https://huggingface.co/datasets/asapp/slue.snips_slu_v1.0
Dataset Card for SNIPS SLU v1.0
Dataset Summary
This dataset contains SNIPS SLU Speech Recognition Dataset, available here.
It contains recordings of commands for smart home appliances in English, with info about demographics of the speaker.
slurp-ear-masked-eval
SLURP: Semantic-Acoustic Masked Evaluation Dataset (EAR Metric)
Dataset Description
This dataset is a custom evaluation subset derived from the SLURP (Spoken Language Understanding Resource Package) dataset.
It is specifically engineered to evaluate the EAR (Execution and Repair) metric for active voice assistants. Using a forced-alignment masking protocol, real human audio is mathematically perturbed with high-amplitude white noise to create two strictly controlled… See the full description on the dataset page: https://huggingface.co/datasets/keylazy/slurp-ear-masked-eval.snips_slu_v1.0
Dataset Card for SNIPS SLU v1.0
Dataset Summary
This dataset contains SNIPS SLU Speech Recognition Dataset, available here.
It contains recordings of commands for smart home appliances in English, with info about demographics of the speaker.
