datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sv-SE-asr-cv
Swedish ASR (Common Voice 22, filtered + rebalanced)
Swedish (sv-SE) speech for ASR, built from Mozilla Common Voice 22.0 (CC0) via
the open fsicoli/common_voice_22_0 mirror. Built to fine-tune
streaming ASR models (e.g. nvidia/nemotron-3.5-asr-streaming-0.6b).
Splits
Split
Clips
Hours
train
22166
26.1
dev
694
0.8
test
1602
2.0
train = the official validated train split + the filtered other bucket + the excess
dev/test speakers: Common Voice's… See the full description on the dataset page: https://huggingface.co/datasets/LokaalHub/sv-SE-asr-cv.midi-svs
[WIP] MIDI SVS
A richly annotated English Suno vocal+MIDI dataset featuring 3k+ curated tracks with stems, transcriptions, lyrics, and structural metadata for SVS music AI and MIR purposes
Attribution
Suno Various 94k
Facebook Demucs
ASLP-lab SongFormer
LinTO AI Whisper Timestamped
ROSVOT
Project Los Angeles
Tegridy Code 2026
