datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
original-songs
Dataset Card for "original-songs" (Audio + análisis DSP)
Dataset Summary
Dataset pequeño de canciones originales creadas con IA, cada una con su WAV,
letra transcrita automáticamente (Whisper) y un análisis DSP completo (tempo,
tonalidad, loudness, features perceptuales) además de detección de contenido
explícito. Pensado para quien quiera mejorar modelos open source: extracción
de features musicales, clasificación de audio, transcripción y moderación de
letras.… See the full description on the dataset page: https://huggingface.co/datasets/arnauquest/original-songs.tkt-song
TKT Song
Public Taiwanese Hokkien song dataset in Hugging Face AudioFolder layout.
This dataset contains rows where all three resources are available and line-aligned:
song audio converted to 16 kHz mono FLAC
Han-character lyrics in lyrics.han
Hokkien romanization in lyrics.hokkien
Current build row count: 29384.
Schema
audio: Hugging Face Audio feature generated from file_name
source_id: source lyrics id
title: source song title
duration_seconds: YouTube… See the full description on the dataset page: https://huggingface.co/datasets/voidful/tkt-song.kazakh_songs_asr
Kazakh Songs ASR Dataset
Dataset Summary
This dataset consists of manually aligned audio–text pairs extracted from Kazakh songs and designed for research in automatic speech recognition (ASR) for low-resource languages. The primary goal of the dataset is to investigate whether sung speech can serve as a complementary training resource for Kazakh ASR systems.
The corpus contains line-level vocal segments obtained from commercially released songs, with manually verified… See the full description on the dataset page: https://huggingface.co/datasets/yeshpanovrustem/kazakh_songs_asr.new_dataset_songs
