speech-disfluency
disfluency_speech_english
Nyra Disfluency Speech English
nyrahealth/disfluency_speech_english is an English speech dataset for evaluating verbatim ASR: models that should transcribe not only the intended words, but also fillers, cutoffs, repetitions, and sound events.
This dataset is based on the AMAAI Lab DisfluencySpeech dataset and reformatted for verbatim-transcription benchmarking with paired:
verbatim_transcript: what the speaker actually said
intended_transcript: a cleaned version of what the… See the full description on the dataset page: https://huggingface.co/datasets/nyralabs/disfluency_speech_english.disfluency_speech_german
Nyra Disfluency Speech German
nyrahealth/disfluency_speech_german is a German speech dataset for evaluating verbatim ASR: models that should transcribe not only the intended words, but also fillers, cutoffs, repetitions, and sound events.
This dataset was recorded in-house by two Nyra researchers, Berns and Laurin, with the goal of producing natural disfluent German speech similar in spirit to the English AMAAI Lab DisfluencySpeech dataset.
Like the English release, it is… See the full description on the dataset page: https://huggingface.co/datasets/nyralabs/disfluency_speech_german.disfluency-speech-normalized
