datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sada_preprocessed_whisper_smallwhisper-small-ko-meeting-datawhisper_transcriptions.reazonspeech.smallatcosim_dataset_for_finetune_whisper_smallti-audio-feature-whisper-smalltir-fidel-whisper-small-datasetwhisper-small-hindi
Dataset Card for "whisper-small-hindi"
More Information needed
arc_whisper_medium_transcriptions.reazonspeech.small
Dataset Card for "arc_whisper_medium_transcriptions.reazonspeech.small"
More Information needed
arc_whisper_transcriptions.reazonspeech.small
Dataset Card for "arc_whisper_transcriptions.reazonspeech.small"
More Information needed
whisper_small_SpNT_infer_result_6.04Meval-whisper-small-english-eka-hard-20260408-1925
Evaluation Results: whisper-small-english
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
leduckhai/MultiMed-ST/asr/whisper-small-english
49.12%
35.11%
Source Data
Evaluation Dataset: Trelis/eka-hard
Model Evaluated: leduckhai/MultiMed-ST/asr/whisper-small-english
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-english-eka-hard-20260408-1925.whisper_transcriptions.reazonspeech.small.wer_10.0eval-whisper-small-eka-hard-20260408-1924
Evaluation Results: whisper-small
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
openai/whisper-small
520.05%
278.24%
Source Data
Evaluation Dataset: Trelis/eka-hard
Model Evaluated: openai/whisper-small
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate for this sample
cer… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-eka-hard-20260408-1924.whisper-small-ru-entity-correctionsarc_whisper_transcriptions.reazonspeech.small.wer_10.0
Dataset Card for "arc_whisper_transcriptions.reazonspeech.small.wer_10.0"
More Information needed
WhisperSmallTest20000
Dataset Card for "WhisperSmallTest20000"
More Information needed
atco2_test_dictation_by_whisper_small_and_llama2_originalarc_whisper_medium_transcriptions.reazonspeech.small.wer_10.0
Dataset Card for "arc_whisper_medium_transcriptions.reazonspeech.small.wer_10.0"
More Information needed
hmong-whisper-smallwhisper_smallWhisperSmallTest200001
Dataset Card for "WhisperSmallTest200001"
More Information needed
WhisperSmallTestmp3
Dataset Card for "WhisperSmallTestmp3"
More Information needed
whisper.kotoba.small.wer_10.0
Dataset Card for "whisper.kotoba.small.wer_10.0"
More Information needed
eval-whisper-small-english-multimed-hard-20260408-1935
Evaluation Results: whisper-small-english
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
leduckhai/MultiMed-ST/asr/whisper-small-english
11.46%
7.46%
Source Data
Evaluation Dataset: Trelis/multimed-hard
Model Evaluated: leduckhai/MultiMed-ST/asr/whisper-small-english
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-english-multimed-hard-20260408-1935.ATCOSIM_dictation_by_finetuned_whisper_smalleval-whisper-small-medical-terms-2025-20260408-1929
Evaluation Results: whisper-small
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
openai/whisper-small
12.95%
4.24%
Source Data
Evaluation Dataset: Trelis/medical-terms-2025
Model Evaluated: openai/whisper-small
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate for this sample… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-medical-terms-2025-20260408-1929.eval-whisper-small-english-medical-terms-2025-20260408-1931
Evaluation Results: whisper-small-english
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
leduckhai/MultiMed-ST/asr/whisper-small-english
13.44%
4.94%
Source Data
Evaluation Dataset: Trelis/medical-terms-2025
Model Evaluated: leduckhai/MultiMed-ST/asr/whisper-small-english
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-english-medical-terms-2025-20260408-1931.WhisperSmallTest200002
Dataset Card for "WhisperSmallTest200002"
More Information needed
whisper-small-hindiwhisper-small-hindi
Dataset Card for "whisper-small-hindi"
More Information needed
