CoolFace
3 results

speech-disfluency

nyralabs /disfluency_speech_english Nyra Disfluency Speech English nyrahealth/disfluency_speech_english is an English speech dataset for evaluating verbatim ASR: models that should transcribe not only the intended words, but also fillers, cutoffs, repetitions, and sound events. This dataset is based on the AMAAI Lab DisfluencySpeech dataset and reformatted for verbatim-transcription benchmarking with paired: verbatim_transcript: what the speaker actually said intended_transcript: a cleaned version of what the… See the full description on the dataset page: https://huggingface.co/datasets/nyralabs/disfluency_speech_english.audioautomatic-speech-recognition1K<n<10K3 likes245 downloads2mo agoHugging Facenyralabs /disfluency_speech_german Nyra Disfluency Speech German nyrahealth/disfluency_speech_german is a German speech dataset for evaluating verbatim ASR: models that should transcribe not only the intended words, but also fillers, cutoffs, repetitions, and sound events. This dataset was recorded in-house by two Nyra researchers, Berns and Laurin, with the goal of producing natural disfluent German speech similar in spirit to the English AMAAI Lab DisfluencySpeech dataset. Like the English release, it is… See the full description on the dataset page: https://huggingface.co/datasets/nyralabs/disfluency_speech_german.audioautomatic-speech-recognitionn<1K2 likes55 downloads2mo agoHugging Facemohammed-bahumaish /disfluency-speech-normalizedgatedaudio1K<n<10K0 likes3 downloads6mo agoHugging Face