CoolFace
Datasetpublic

zbrunner/speakeroverlap_multiseg

MultiSeg Dataset for ASR Hallucinations Description MultiSeg is a perturbed and altered version of the TEDLIUM3 dataset, specifically created for evaluating the robustness of Automatic Speech Recognition (ASR) systems. This dataset is derived from the 'speakeroverlap' subset, which consists of held-back training data from TEDLIUM3. Purpose The primary purpose of the MultiSeg dataset is to: Elicit hallucinations from ASR systems Evaluate ASR… See the full description on the dataset page: https://huggingface.co/datasets/zbrunner/speakeroverlap_multiseg.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes7downloads
5 commits on main
5b4373d2y ago

Update README.md

zbrunner
f1ba3ad2y ago

Update README.md

zbrunner
24a783e2y ago

Update README.md

zbrunner
5313ee32y ago

Upload folder using huggingface_hub

zbrunner
bd697032y ago

initial commit

zbrunner