datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
neuro-parakeet-food
neuro-whisper-v1
Dataset Description
This is a synthetic dataset for German medical speech recognition, specifically designed for fine-tuning ASR models on neuro-oncology and neurology terminology. The dataset provides a comprehensive coverage of German medical terminology in the neurology and neuro-oncology domains.
Data Generation
Voice Data: Synthetically generated using Resemble AI Chatterbox TTS
Text Data: Medical text generated with Qwen/Qwen3-30B-A3B… See the full description on the dataset page: https://huggingface.co/datasets/NeurologyAI/neuro-parakeet-food.J-HARD-TTS-Eval
J-HARD-TTS-Eval
[!NOTE]
For full documentation, detailed benchmark results, and methodology, please refer to the GitHub Repository.
Overview
J-HARD-TTS-Eval is a benchmark designed to evaluate the robustness of autoregressive Japanese Text-To-Speech (TTS) models.
It focuses on specific failure modes such as stability in short sequences, repetition handling, and context completion.
Usage
You can easily load the dataset using the Hugging Face datasets… See the full description on the dataset page: https://huggingface.co/datasets/Parakeet-Inc/J-HARD-TTS-Eval.parakeet-tdt-blind-spots
Blind Spots of nvidia/parakeet-tdt-0.6b-v2
This dataset documents 14 systematically identified blind spots in NVIDIA's parakeet-tdt-0.6b-v2 automatic speech recognition model. The errors span 8 distinct categories and reveal a consistent pattern: the model struggles with inputs outside the distribution of its Western English-centric training data.
Model Under Test
Property
Value
Model
nvidia/parakeet-tdt-0.6b-v2
Parameters
600M
Architecture… See the full description on the dataset page: https://huggingface.co/datasets/TieIncred/parakeet-tdt-blind-spots.eval-parakeet-tdt-0.6b-v3-medical-terms-2025-20260408-1926
Evaluation Results: parakeet-tdt-0.6b-v3
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
nvidia/parakeet-tdt-0.6b-v3
11.34%
3.63%
Source Data
Evaluation Dataset: Trelis/medical-terms-2025
Model Evaluated: nvidia/parakeet-tdt-0.6b-v3
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-parakeet-tdt-0.6b-v3-medical-terms-2025-20260408-1926.eval-parakeet-tdt-0.6b-v3-eka-hard-20260408-1920
Evaluation Results: parakeet-tdt-0.6b-v3
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
nvidia/parakeet-tdt-0.6b-v3
37.59%
20.64%
Source Data
Evaluation Dataset: Trelis/eka-hard
Model Evaluated: nvidia/parakeet-tdt-0.6b-v3
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate for… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-parakeet-tdt-0.6b-v3-eka-hard-20260408-1920.parakeet-stt-redone
parakeet-stt-redone
What this is
108,276 raw→clean transcript pairs sourced from
aldigobbler/stt-correction,
re-labeled using GLM-5.1-FP8 as the teacher model with our production
cleanup prompt.
How it differs from the source dataset
aldigobbler/stt-correction
this dataset
Target
Verbatim transcript restoration (lowercase, no punctuation, fillers kept/restored)
Polished readable text — punctuated, paragraphed, fillers selectively removed… See the full description on the dataset page: https://huggingface.co/datasets/rdsm/parakeet-stt-redone.eval-parakeet-tdt-0.6b-v3-multimed-hard-20260408-1930
Evaluation Results: parakeet-tdt-0.6b-v3
Evaluation results from Whisper model evaluation.
Summary
Model
WER
CER
nvidia/parakeet-tdt-0.6b-v3
15.94%
10.13%
Source Data
Evaluation Dataset: Trelis/multimed-hard
Model Evaluated: nvidia/parakeet-tdt-0.6b-v3
Columns
Column
Description
audio
Audio sample (if available from source dataset)
reference
Ground truth transcription
prediction
Model prediction
wer
Word Error Rate… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-parakeet-tdt-0.6b-v3-multimed-hard-20260408-1930.moe-speech-plus-cache-parakeet-v1-audit
MoeSpeechPlus Parakeet cache v1 — aggregate audit
This aggregate-only audit describes
otoha-project/moe-speech-plus-cache-parakeet-v1 at revision
fa80a0e57e4af06df434130f66a2792207f00c45. It excludes speaker IDs, utterance IDs,
paths, text, tensors, pickle data, and cache files.
The two tar parts form a continuous stream and terminate correctly. The repository also
contains a 31,232-file expanded tree; one sampled expanded file was byte-identical to its
tar member. The cache… See the full description on the dataset page: https://huggingface.co/datasets/otoha-project/moe-speech-plus-cache-parakeet-v1-audit.mmm_project_parakeetmmm_project_parakeet_intermdata
