CoolFace
16 results

long_audio

mitermix /audiosnippets_long_2_8M2 likes4.2k downloads2y agoHugging Faceliumindmind /Neko_Audio-30K_Longaudio10K<n<100K7 likes3k downloads4mo agoHugging Facemitermix /audiosnippets_long_2_5Maudio1M<n<10M3 likes1.4k downloads2y agoHugging Faceholvan /LongAudioSpan LongAudioSpan: Spanning the Duration and Depth of Audio Comprehension Introduction LongAudioSpan is a benchmark for long-form audio comprehension, spanning diverse durations and cognitive depths. Questions come from two complementary paths: Native QA: questions drawn from the audio's natural content. Anchor QA: questions built around acoustic anchors planted into the audio. Each path is scored in its own mode: Accuracy: multiple choice… See the full description on the dataset page: https://huggingface.co/datasets/holvan/LongAudioSpan.textaudio-text-to-text1K<n<10K9 likes957 downloads7d agoHugging Facemitermix /audiosnippets_long_1Maudio100K<n<1M0 likes822 downloads2y agoHugging Faceai-music4you3 /enhanced-audiosnippets-long-2-8M Enhanced Audiosnippets Long 2.8M Enhanced version of mitermix/audiosnippets_long_2_8M with speech enhancement, emotion annotations, speaker embeddings, and comprehensive metadata analysis. Dataset Summary Metric Value Total samples 2,633,037 Total audio hours 4,932 h Duration range 3.0s - 1124.3s Mean duration 6.7s Audio format WAV, 48kHz mono Tar files 1,410 Processing Pipeline Each audio sample was processed through: Speech… See the full description on the dataset page: https://huggingface.co/datasets/ai-music4you3/enhanced-audiosnippets-long-2-8M.tabularaudio-classification1M<n<10M1 likes429 downloads6mo agoHugging Face