datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
random-audiosor_in_datasettimbre_range
Dataset Card for Timbre and Range Dataset
Dataset Summary
The timbre dataset contains acapella singing audio of 9 singers, as well as cut single-note audio, totaling 775 clips (.wav format)
The vocal range dataset includes several up and down chromatic scales audio clips of several vocals, as well as the cut single-note audio clips (.wav format).
Supported Tasks and Leaderboards
Audio classification
Languages
Chinese, English
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/ccmusic-database/timbre_range.oai-5g-srs-ranging-dataset
OAI 5G NR SRS Ranging Captures
Uplink Sounding Reference Signal (SRS) channel-estimate captures from a monolithic 5G NR
software-defined-radio testbed, collected for SRS-based ranging experiments.
The gNB (OpenAirInterface on a USRP X410, Band n78, 40 MHz / 106 PRB) configures each UE
to transmit SRS; the gNB's per-SRS frequency-domain channel estimate, oversampled IDFT CIR,
and ToA estimate are streamed off the PHY via OAI's T_tracer and recorded at a series of
known… See the full description on the dataset page: https://huggingface.co/datasets/ahancock516/oai-5g-srs-ranging-dataset.random_imgrandom-chunks-dsSinhalaASR-testindic_voices_hindi_only_plus_vaani_random_sample_34548_4616_seed_43_cleanArabic_Audio_Deepfake
ArAD Dataset (Arabic Audio DeepFake Dataset)
Dataset SummaryThis dataset contains Arabic deepfake audio samples, focusing mainly on Levantine dialect with some examples in Standard Arabic. It was created using the RVC v2 framework, fine-tuned on a custom dataset of multi-dialect Arabic speech. The goal is to simulate real-world deepfake audio attacks by generating synthetic speech from actual recordings and voice messages.
One of the first datasets to include real-world deepfake… See the full description on the dataset page: https://huggingface.co/datasets/DeepFake-Audio-Rangers/Arabic_Audio_Deepfake.SinhalaASR-1000indic_voices_hindi_only_random_sample_17274_2308_seed_42indic_voices_hindi_only_random_sample_17274_2308_seed_42_cleantamil-english-podcast-diarization
Tamil-English Code-Mixed Podcast Diarization Dataset
Dataset Summary
This dataset contains long-form Tamil-English code-mixed podcast recordings
annotated for speaker diarization research. The recordings consist of natural
conversational speech with multiple speakers and realistic acoustic conditions,
making the dataset suitable for evaluating diarization pipelines in
real-world scenarios.
The dataset is intended to support research in:
Speaker diarization
Code-mixed… See the full description on the dataset page: https://huggingface.co/datasets/Rangasuthan/tamil-english-podcast-diarization.indic_voices_hindi_only_random_sample_17334_2312_seed_42indic_voices_hindi_only_plus_vaani_random_sample_17274_2308_seed_43_cleanEgyptian_TTS3RSSaudiTalk
SaudiTalk: A Multi-Source Dialectal Speech Dataset from Saudi Arabia
Dataset Summary
SaudiTalk is a curated and human-verified Arabic speech dataset covering three major Saudi dialects: Hijazi, Ha’il, and Southern. The dataset is constructed from publicly available social media content and is designed to support research in Automatic Speech Recognition (ASR), dialect identification, and Arabic speech processing.
Key Features
3 Saudi dialects:… See the full description on the dataset page: https://huggingface.co/datasets/ranaRan689/SaudiTalk.lili-romanian-single-speaker-piper-cleannedu_randyPeterdisplace_dev_datasbbdata_snr_random
Dataset Card for "sbbdata_snr_random"
More Information needed
eval_bell_ring_put_tape_in_bin_random_init_testThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101_follower",
"total_episodes": 12,
"total_frames": 6383,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 500,
"audio_files_size_in_mb": 100,
"fps": 30,
"splits": {
"train": "0:12"
},
"data_path":… See the full description on the dataset page: https://huggingface.co/datasets/AdamMalyshev/eval_bell_ring_put_tape_in_bin_random_init_test.radasrandom_dataflock-demo-automatic-speech-recognition-sectionsBadasranveer_audiodatasetcorpus5-proposed-no-overlap-random-1000
Corpus5 proposed-rule unflagged random sample
This manually gated dataset contains 1,000 uniformly randomly selected rows from
the fixed 2,715,793-row prepared Corpus5 snapshot on mac02.
A row is eligible only when both proposed checks are false:
the full chunk interval does not intersect positive-duration diarization turns
from two distinct speakers; and
no speaker-change/no-change disagreement is detected at aligned adjacent words
in transcript1 and transcript2 and propagated… See the full description on the dataset page: https://huggingface.co/datasets/Cybrpgs/corpus5-proposed-no-overlap-random-1000.L2EnglishFluency_speechocean762-RankingGAFAS
