datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
HyperspectralCollections
Hyperspectral Image Collections
Collect the open-source hyperspectral images and train any deep models.
wikisumIALP-2026-data
IALP-2026: Whisper Open-Set Data-Selection — Query / Dev / Test Sets
Supporting data for the study "Whisper-Based Open-Set Data Selection for NSC
Adaptation." This repository holds the fixed target-query, validation, and
evaluation sets used across all experiments. Each part is a self-contained
.tar.gz.
All audio is 16 kHz mono. Each split ships with:
audio/ — audio files (FLAC, except GigaSpeech which is WAV PCM_16)
wav.scp — <utt_id> audio/<file> (Kaldi-style, relative paths)… See the full description on the dataset page: https://huggingface.co/datasets/pengyizhou/IALP-2026-data.octo_iamlab_cmu_pickup_insertosu-beatmaps-duplicated
osu! Beatmaps Dataset (WebDataset)
A collection of ranked/loved osu! beatmaps with audio and chart data, in WebDataset format.
Dataset Variants
Variant
Audio Format
Description
original
MP3/OGG/WAV
Full quality original audio files
compressed
64kbps Mono Opus
Compressed audio for smaller download
from datasets import load_dataset
# Load original audio variant
ds = load_dataset("project-riz/osu-beatmaps", "original", streaming=True)
# Load compressed… See the full description on the dataset page: https://huggingface.co/datasets/IamXiangyu/osu-beatmaps-duplicated.pt-br-tts-iasmin-qwen3
pt-br-tts-iasmin-qwen3
21957 clips PT-BR sintetizados com Qwen3-TTS-12Hz-1.7B. Voz Iasmin (voice-clone, is_iasmin=true, ~13957 clips) + vozes diversas CustomVoice (Ryan, Aiden, Vivian, Dylan, is_iasmin=false). WAV em tar shards (WebDataset); transcricao, voz, is_iasmin e sr em metadata.jsonl.
mult-dataIASA
