datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
audio-fingerprint-indian-bench
Audio Fingerprinting Benchmark on Indian Classical Music
A reproducible, pre-registered benchmark of five audio-fingerprinting systems on the Saraga 1.5 corpus (Hindustani + Carnatic), plus a pre-registered training-recipe improvement to the NAFP baseline that achieves Bonferroni-significant gains on 1-second queries.
v0.7 · closed-world retrieval · 5 systems · 6 528 evaluation cells · pooled McNemar p = 3.18 × 10⁻⁶
TL;DR
5 systems benchmarked: Olaf, Dejavu… See the full description on the dataset page: https://huggingface.co/datasets/Tachyeon/audio-fingerprint-indian-bench.audio-fingerprint-indian-bench
Audio Fingerprinting Benchmark on Indian Classical Music
A reproducible, pre-registered benchmark of five audio-fingerprinting systems on the Saraga 1.5 corpus (Hindustani + Carnatic), plus a pre-registered training-recipe improvement to the NAFP baseline that achieves Bonferroni-significant gains on 1-second queries.
v0.7 · closed-world retrieval · 5 systems · 6 528 evaluation cells · pooled McNemar p = 3.18 × 10⁻⁶
TL;DR
5 systems benchmarked: Olaf, Dejavu… See the full description on the dataset page: https://huggingface.co/datasets/aryanBanwala/audio-fingerprint-indian-bench.music-fingerprint-dataset
Neural Audio Fingerprint Dataset
(c) 2021 by Sungkyun Chang
https://github.com/mimbres/neural-audio-fp
This dataset includes all music sources, background noise and impulse-reponses
(IR) samples that have been used in the work ["Neural Audio Fingerprint for
High-specific Audio Retrieval based on Contrastive Learning"]
(https://arxiv.org/abs/2010.11910).
Format:
16-bit PCM Mono WAV, Sampling rate 8000 Hz
Description:
/
fingerprint_dataset_icassp2021/… See the full description on the dataset page: https://huggingface.co/datasets/arch-raven/music-fingerprint-dataset.voicebank-demand-fingerprint-16k
Enhanced VoiceBank+DEMAND Dataset with Noise Fingerprints (16000Hz)
Dataset Description
⚠️ Important Note: This is NOT the exact same dataset as the original VoiceBank-DEMAND dataset. While we follow the VoiceBank-DEMAND specification for speaker splits and SNR levels, the noise types used are different due to DEMAND dataset availability and environmental focus.
This dataset contains enhanced noisy speech samples created by mixing clean speech from the VCTK corpus with… See the full description on the dataset page: https://huggingface.co/datasets/yairamr/voicebank-demand-fingerprint-16k.voicebank-demand-fingerprint-48k
Enhanced VoiceBank+DEMAND Dataset with Noise Fingerprints (48000Hz)
Dataset Description
⚠️ Important Note: This is NOT the exact same dataset as the original VoiceBank-DEMAND dataset. While we follow the VoiceBank-DEMAND specification for speaker splits and SNR levels, the noise types used are different due to DEMAND dataset availability and environmental focus.
This dataset contains enhanced noisy speech samples created by mixing clean speech from the VCTK corpus with… See the full description on the dataset page: https://huggingface.co/datasets/yairamr/voicebank-demand-fingerprint-48k.audio-fingerprint-db
