CoolFace
3 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nyralabs /disfluency_speech_english Nyra Disfluency Speech English nyrahealth/disfluency_speech_english is an English speech dataset for evaluating verbatim ASR: models that should transcribe not only the intended words, but also fillers, cutoffs, repetitions, and sound events. This dataset is based on the AMAAI Lab DisfluencySpeech dataset and reformatted for verbatim-transcription benchmarking with paired: verbatim_transcript: what the speaker actually said intended_transcript: a cleaned version of what the… See the full description on the dataset page: https://huggingface.co/datasets/nyralabs/disfluency_speech_english.audioautomatic-speech-recognition1K<n<10K3 likes247 downloads2mo agoHugging Face02nyralabs /disfluency_speech_german Nyra Disfluency Speech German nyrahealth/disfluency_speech_german is a German speech dataset for evaluating verbatim ASR: models that should transcribe not only the intended words, but also fillers, cutoffs, repetitions, and sound events. This dataset was recorded in-house by two Nyra researchers, Berns and Laurin, with the goal of producing natural disfluent German speech similar in spirit to the English AMAAI Lab DisfluencySpeech dataset. Like the English release, it is… See the full description on the dataset page: https://huggingface.co/datasets/nyralabs/disfluency_speech_german.audioautomatic-speech-recognitionn<1K2 likes53 downloads2mo agoHugging Face03mohammed-bahumaish /disfluency-speech-normalizedgatedaudio1K<n<10K0 likes3 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.