datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SpotSound-Bench
SpotSound-Bench: A 'Needle-in-a-Haystack' Evaluation for Audio Temporal Grounding
Benchmark Summary
SpotSound-Bench is a challenging temporal grounding benchmark designed to evaluate Large Audio-Language Models (ALMs).
Existing benchmarks for audio temporal grounding often feature high ratios of target-window duration to full audio clip duration, which fail to simulate real-world scenarios where short events are obscured by dense background sounds. To bridge… See the full description on the dataset page: https://huggingface.co/datasets/Loie/SpotSound-Bench.podcasts_spotifyhubert_process_filter_spotifyhub_large_ls960-ft_spot_data_allhubert_processed_spotifywhisp_base_spot_data_allwhisp_tiny_spot_data_allspot_data_womenspot_data_allspotify_musicwav2vec2_phoneme_spot_data_allspotify_dataset_whisperspotify-20k-testraw_train_dev_test_spotifyspotify_raw_small_2parakeet-tdt-blind-spots
Blind Spots of nvidia/parakeet-tdt-0.6b-v2
This dataset documents 14 systematically identified blind spots in NVIDIA's parakeet-tdt-0.6b-v2 automatic speech recognition model. The errors span 8 distinct categories and reveal a consistent pattern: the model struggles with inputs outside the distribution of its Western English-centric training data.
Model Under Test
Property
Value
Model
nvidia/parakeet-tdt-0.6b-v2
Parameters
600M
Architecture… See the full description on the dataset page: https://huggingface.co/datasets/TieIncred/parakeet-tdt-blind-spots.wav2vec_filter_wo_processing_spotifyatc-spotlighthebrew_keyword_spotspots_audios
Dataset Card for "spots_audios"
More Information needed
spot_data_spanglishspotify_raw_small_1spot_data_aave_menKeyword_Spottingwhisp_small_spot_data_allhub_base_ls960-ft_spot_data_allspot_data_menwhisp_med_spot_data_allhubert_small1_filter_allcolumns_spotifyspot_data_aave_women
