datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
low_quality_call_voice
Dataset Card for "low_quality_call_voice"
More Information needed
low-germanLow-Language-TTS
Low-Language-TTS
A held-out TTS/ASR test set for the long tail of
malaysia-ai/Multilingual-TTS:
50 of the lowest-resource languages in that corpus,
25 utterances each (1250 rows, 2.56 hours).
Every row carries the three things needed to score a NeuCodec speech-token model
without touching the parent corpus:
column
what
audio
the original clip, exactly as stored upstream (mostly mp3)
tokens
NeuCodec speech tokens at 50 tokens/s — the <|s_N|> ids the Multilingual-TTS… See the full description on the dataset page: https://huggingface.co/datasets/malaysia-ai/Low-Language-TTS.sukasuka-anime-vocal-datasetthis dataset is the parquet version of the dataset that was created by mio
original dataset link : https://huggingface.co/datasets/mio/sukasuka-anime-vocal-dataset
please make sure to follow and heart react the original author (≧∇≦)ノ
uaspeech_tts_very_lowvertin_train_datasetAdaption-low-resource-audio
Adaption Low-Resource Audio
A low-resource-language subset of
Reubencf/PolyglotAudio,
remastered with Adaption's Adaptive Data
platform. Each row carries the original Tatoeba-derived audio clip
alongside sharpened enhanced_prompt / enhanced_completion columns
so the data is ready for speech-model fine-tuning and evaluation on
languages that are typically under-represented in open ASR/TTS corpora.
Dataset size
3,704 rows of paired audio + text, spanning 10 languages… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/Adaption-low-resource-audio.alsallom_update_para_UAE_transcription_low_chunk_by_elevenlabhigh-sound-and-low-musicinterleaving_mmau_lowest_pitchhmong_low_bleuuaspeech_tts_lowlow-decoder
low-decoding
Author: Surpem
This dataset contains 1200 unique, clean synthetic audio signals representing decoded text commands.
The audio signals represent synthesized Morse Code message blocks.
Dataset Structure
id: A unique UUID string.
audio: The audio wav bytes (16kHz Mono).
text: The decoded string transcription.
Dataset Level: LOW
Low: Slow WPM (~12 WPM), high signal-to-noise ratio (clean), short command strings.
Medium: Fast WPM (~24 WPM), background… See the full description on the dataset page: https://huggingface.co/datasets/Surpem/low-decoder.43-143-phase2-appconv-ime-low_pitchLaila_low_pauselow_quality_khmer_speechDubeningVietMuong-LowResourceLow_Resource_Arabic_Adaptationspeech_mmau_lowest_pitchstt_lowrank_finetuningeval-whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten-20260223-2259
Training Evaluation: whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten
Evaluation results comparing base model vs fine-tuned model.
Summary
Model
WER
openai/whisper-small (base)
46.81%
Trelis/whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten (fine-tuned)
34.22%
Improvement: 12.59% WER reduction (lower is better)
Source Data
Evaluation Dataset: Trelis/pilotgpt-test-0.5s-rewritten
Base Model: openai/whisper-small… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten-20260223-2259.pilotgpt-unified-all-data-lowercase-new-rewrittenpilotgpt-unified-all-data-lowercase-data-prepeval-whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772-20260219-1448
Training Evaluation: whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772
Evaluation results comparing base model vs fine-tuned model.
Summary
Model
WER
openai/whisper-small (base)
53.69%
Trelis/whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772 (fine-tuned)
27.54%
Improvement: 26.15% WER reduction (lower is better)
Source Data
Evaluation Dataset: Trelis/pilotgpt-test-0.5s
Base Model: openai/whisper-small… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772-20260219-1448.SimpleScript_HouseNumberSpeechgenDataset_lowercase
