datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Whisper-fine-tune-2uyghur-whisper-finetune
UyZh-FolkSpeech Whisper 微调数据集
维吾尔语-汉语平行语音数据集,适用于 Whisper 模型微调。
数据集来源
本数据集源自 UyZh-FolkSpeech,经过以下处理:
音频格式转换: M4A → WAV (16kHz, 单声道)
文本清洗: 移除不可见控制字符 (U+200E, U+200F 等)
数据集划分: train/val/test (80%/10%/10%)
目录重组: 按划分分文件夹存储
数据统计
划分
记录数
音频数
时长
train
6,348
6,348
407.60 分钟
validation
793
793
48.89 分钟
test
795
795
51.80 分钟
总计
7,936
7,936
508.29 分钟
内容分布
短句 (sentence): 3,812 条
词汇短语 (word): 4,124 条
说话人分布
每个文本条目由 4 位说话人… See the full description on the dataset page: https://huggingface.co/datasets/anke01/uyghur-whisper-finetune.Whisper-fine-tune-1atcosim_dataset_for_finetune_whisper_smallwhisper-finetune-audio_test6
Whisper Fine-tuning Dataset
This dataset contains 30-second chunks of audio and corresponding transcripts for fine-tuning OpenAI Whisper models.
Each JSON entry maps a .wav file (16kHz mono) to its transcription.
feji-first-finetune
FEJI First Fine-Tune
FEJI First Fine-Tune is a Turkish folk music dataset prepared for ACE-Step 1.5
LoRA fine-tuning experiments. It contains 201 audio examples with aligned
ACE-Step metadata for caption-conditioned music generation.
The dataset was exported from the local finetune-dataset/ folder and uploaded
as Parquet shards with embedded audio. Each row contains one audio sample plus
metadata fields used by the ACE-Step training and dataset-builder workflow.… See the full description on the dataset page: https://huggingface.co/datasets/alibayram/feji-first-finetune.whisper-finetune-audio_test5basebend_finetuneWhisper-Fine-Tune-One-Shot-Eval
Whisper Fine-Tuning Evaluation: Local vs Commercial ASR
A "back of the envelope" evaluation comparing fine-tuned Whisper models running locally against commercial ASR APIs via Eden AI.
The Question
Can fine-tuning Whisper achieve measurable WER reductions, even when comparing local inference against cloud-based commercial models?
TL;DR
Yes. Fine-tuned Whisper Large Turbo running locally achieved 5.84% WER, beating the best commercial API (Assembly at… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/Whisper-Fine-Tune-One-Shot-Eval.orpheus-finetune-datasetfinal-finetune-asr-datasetorpheus_finetune_dataset2026_08_14_2325__cosyvoice_flow_finetuned_on__abacus__sentence_ch__from_baseFineTune_Kannadamllm_finetuneWavLM-base-dataset_finetuneDataset_overlap_dan_tidak_overlap_finetuneindicf5-nfe-finetune-testarabic-finetuneWhisper-fine-tune-p-23tts-finetuneWhisper-fine-tune-p-14fine-tune-argos-0.5finetuned_avg_pooling_DF_Audio_Embeddingslibrispeech_small_asr_fine-tunefine-tune-hebrew-dataset
Dataset Card for "fine-tune-hebrew-dataset"
More Information needed
diar-finetune-dataset-chatterbox-dialoguesUmkehr-fine-tunefine-tune-hebrew-dataset-2
Dataset Card for "fine-tune-hebrew-dataset-2"
More Information needed
in_context_QA_ASR_TTS_finetune_3-2-11B_rank64_ls960_replay_v4
