datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
synthesized_audio
Dataset Card for "synthesized_audio"
More Information needed
waxal-autolabled
Auot-Lableing Waxal unlabeled dataset on Best Multilingual Ethio-ASR models
@article{abdullah2026ethio,
title={Ethio-ASR: Joint Multilingual Speech Recognition and Language Identification for Ethiopian Languages},
author={Abdullah, Badr M and Azime, Israel Abebe and Tonja, Atnafu Lambebo and Alabi, Jesujoba O and Alemu, Abel Mulat and Hagos, Eyob G and Balcha, Bontu Fufa and Nerea, Mulubrhan A and Yadeta, Debela Desalegn and Marilign, Dagnachew Mekonnen and others}… See the full description on the dataset page: https://huggingface.co/datasets/israel/waxal-autolabled.ftarsynthesized_audio_part2auto-pale
Dataset card for pale
Dataset summary
This dataset contains league of legends champions' quotes parsed from fandom.
See dataset usage example at google colab.
The dataset is available in the following configurations:
vanilla - all data pulled from the website without significant modifications apart from the web page structure parsing;
quotes - truncated version of the corpus, which does't contain sound effects;
annotated - an extended version of the full configuration… See the full description on the dataset page: https://huggingface.co/datasets/zeio/auto-pale.auto-batch
Dataset Card for "auto-batch"
More Information needed
auto-whisper-disfluencymtg_jamendo_autotagging
🎵 MTG-Jamendo Autotagging (30s, 16kHz, Multi-Label)
This dataset is a curated subset of the MTG-Jamendo Autotagging Dataset, containing only tracks that include instrument, genre, and mood/theme annotations. Each audio file is preprocessed to ensure consistent formatting for music auto-tagging tasks.
🧾 Dataset Description
Source: MTG-Jamendo Autotagging benchmark
Selection: Tracks that include all three tag types:
genre
instrument
mood/theme
Preprocessing:… See the full description on the dataset page: https://huggingface.co/datasets/vtsouval/mtg_jamendo_autotagging.Auto_Engine_Classification
Engine Sound Windows (YouTube-derived, metadata-only)
Timestamps and weak labels for training an engine-configuration audio classifier (v-twin vs.
inline-4 vs. flat-6, etc.) from short audio windows. This dataset does not contain audio.
Each row points at a public YouTube video id plus a (start_sec, end_sec) window; you fetch
and slice the audio yourself (see Reconstructing audio below).
Why metadata-only
The source audio was collected by searching YouTube (via… See the full description on the dataset page: https://huggingface.co/datasets/joakes90/Auto_Engine_Classification.AutomaticSpeechRecognition_LibriSpeech-TestOther
Dataset Card for "AutomaticSpeechRecognition_LibriSpeech-TestOther"
More Information needed
whisper_auto__audioAutomaticSpeechRecognition_LibriSpeech-TestClean
Dataset Card for "AutomaticSpeechRecognition_LibriSpeech-TestClean"
More Information needed
AutomaticSpeechRecognition_LJSpeech
Dataset Card for "AutomaticSpeechRecognition_LJSpeech"
More Information needed
flock-demo-automatic-speech-recognition-sectionsauto_datasetautoridadeheybee-probe-20260521121507-ad05daef-auto-converted-audio-publicheybee-probe-20260521180915-18e95535-auto-converted-audio-publicautorecord_pw733_230726autolyrics-datasetautomatic_extraction_data_20250115automatic_extraction_data_20250120heybee-probe-20260521121507-ad05daef-auto-converted-audio-gatedheybee-probe-20260521180915-18e95535-auto-converted-audio-gated
