CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ekacare /eka-medical-asr-evaluation-dataset Eka Medical ASR Evaluation Dataset Dataset Overview and Sourcing The Eka Medical ASR Evaluation Dataset enables comprehensive evaluation of automatic speech recognition systems designed to transcribe medical speech into accurate text—a fundamental component of any medical scribe system. This dataset captures the unique challenges of processing medical terminology, particularly branded drugs, which is specific to the Indian context. The dataset comprises over 3,900+… See the full description on the dataset page: https://huggingface.co/datasets/ekacare/eka-medical-asr-evaluation-dataset.audioautomatic-speech-recognition1K<n<10K17 likes1.3k downloads1y agoHugging Face02ekacare /denoising-impact-evaluation-dataset Denoising Impact Evaluation Dataset Dataset Description The ekacare/denoising-impact-evaluation-dataset is a comprehensive benchmark dataset designed to evaluate the effects of speech enhancement on automatic speech recognition (ASR) systems in medical speech contexts. It includes paired noisy and denoised audio subsets under controlled acoustic conditions to support systematic analysis of denoising performance. Source Data Base Dataset:… See the full description on the dataset page: https://huggingface.co/datasets/ekacare/denoising-impact-evaluation-dataset.audiotext-to-speech10K<n<100K0 likes363 downloads9mo agoHugging Face03Scicom-intl /Evaluation-Multilingual-VC Evaluation-Multilingual-VC We use dataset https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon, Filter languages that support by Whisper Large V3 to evaluate WER automatically, Only take test set, sort by up votes. Because VC required to source text, source audio, target text, we make sure the target text is not same as source text, target text we take from other rows. Only build first 500 rows for each language Github issue at… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/Evaluation-Multilingual-VC.audio10K<n<100K0 likes257 downloads6mo agoHugging Face04plnguyen2908 /AudioVisual-Benchmark-Evaluation AudioVisual Benchmark Evaluation — evaluation subsets Item-id lists for the audio-visual benchmark subsets used in our reported evaluation tables. Layout <benchmark>/eval_subset.csv item ids evaluated in the paper <benchmark>/media_index.csv id -> media filename(s) <benchmark>/media/ the media files those ids refer to eval_subset.csv holds a single id column keyed to the source benchmark (question_id, idx, or index). media/ contains exactly the… See the full description on the dataset page: https://huggingface.co/datasets/plnguyen2908/AudioVisual-Benchmark-Evaluation.audiomultiple-choice10K<n<100K0 likes212 downloads24d agoHugging Face05drizzymedia /Stems-Evaluation-Kit 🎧 SonicSets High-Fidelity Stems Evaluation Kit This is a premium evaluation subset provided by SonicSets, the industrial-grade audio data infrastructure for Large Audio Models (LAM). 📊 Dataset Specifications Format: 48kHz / 24-bit Uncompressed WAV (Studio-Grade Ground Truth) Feature: Absolute zero-crosstalk multi-track isolation Environment: Strict anechoic capture (RT60 < 0.2s) Purpose: Fully optimized for training and benchmarking state-of-the-art Source… See the full description on the dataset page: https://huggingface.co/datasets/drizzymedia/Stems-Evaluation-Kit.audion<1K0 likes116 downloads2mo agoHugging Face06sujalappa /nvidia-brain-noise-evaluation-dataset Nvidia Brain Noise Evaluation Dataset Dataset Description This dataset contains 64 samples organized across multiple splits and 32 subsets. The dataset includes audio data. Dataset Structure Subsets This dataset includes the following subsets: noisy-bg-snr-10: 2 samples test: 2 samples noisy-bg-snr-20: 2 samples test: 2 samples noisy-bg-snr-30: 2 samples test: 2 samples noisy-bg-snr-40: 2 samples test: 2 samples noisy-bg-snr-50: 2 samples… See the full description on the dataset page: https://huggingface.co/datasets/sujalappa/nvidia-brain-noise-evaluation-dataset.audioautomatic-speech-recognitionn<1K0 likes95 downloads1y agoHugging Face07deepsafe /evaluation-datasetgated DeepSafe Evaluation Dataset Evaluation set for DeepSafe, a deepfake detection benchmark. Tiers Tier Samples Generators Size Use master_eval_small/ 198 116 1.7 GB smoke test, under 2 min master_eval/ 15,454 411 10 GB the standard benchmark master_eval_full/ 45,954 411 25 GB complete set Medium tier composition: 9,954 image, 3,500 audio, 2,000 video. from huggingface_hub import snapshot_download snapshot_download("deepsafe/evaluation-dataset"… See the full description on the dataset page: https://huggingface.co/datasets/deepsafe/evaluation-dataset.audioaudio-classification10K<n<100K0 likes75 downloads5d agoHugging Face08KothapalliAnusha /eka-medical-asr-evaluation-dataset Eka Medical ASR Evaluation Dataset Dataset Overview and Sourcing The Eka Medical ASR Evaluation Dataset enables comprehensive evaluation of automatic speech recognition systems designed to transcribe medical speech into accurate text—a fundamental component of any medical scribe system. This dataset captures the unique challenges of processing medical terminology, particularly branded drugs, which is specific to the Indian context. The dataset comprises over 3,900+… See the full description on the dataset page: https://huggingface.co/datasets/KothapalliAnusha/eka-medical-asr-evaluation-dataset.audioautomatic-speech-recognition1K<n<10K0 likes71 downloads8mo agoHugging Face09Noothi /telugu-indicf5-evaluationaudion<1K0 likes65 downloads2mo agoHugging Face10sonicsets-data /Stems-Evaluation-Kit 🎧 SonicSets High-Fidelity Stems Evaluation Kit This is a premium evaluation subset provided by SonicSets, the industrial-grade audio data infrastructure for Large Audio Models (LAM). 📊 Dataset Specifications Format: 48kHz / 24-bit Uncompressed WAV (Studio-Grade Ground Truth) Feature: Absolute zero-crosstalk multi-track isolation Environment: Strict anechoic capture (RT60 < 0.2s) Purpose: Fully optimized for training and benchmarking state-of-the-art Source Separation… See the full description on the dataset page: https://huggingface.co/datasets/sonicsets-data/Stems-Evaluation-Kit.audion<1K0 likes51 downloads6mo agoHugging Face11havahavai /eka-medical-asr-evaluation-dataset Eka Medical ASR Evaluation Dataset Dataset Overview and Sourcing The Eka Medical ASR Evaluation Dataset enables comprehensive evaluation of automatic speech recognition systems designed to transcribe medical speech into accurate text—a fundamental component of any medical scribe system. This dataset captures the unique challenges of processing medical terminology, particularly branded drugs, which is specific to the Indian context. The dataset comprises over 3… See the full description on the dataset page: https://huggingface.co/datasets/havahavai/eka-medical-asr-evaluation-dataset.audioautomatic-speech-recognition1K<n<10K1 likes50 downloads4mo agoHugging Face12BatSilver /recorded_evaluation_dataset2audion<1K0 likes44 downloads10mo agoHugging Face13sujalappa /denoised-evaluation-dataset Denoised Subset Fixed Dataset Description This dataset contains 500 samples organized across multiple splits and 1 subsets. The dataset includes audio data. Dataset Structure Subsets This dataset includes the following subsets: denoised: 500 samples test: 500 samples Usage Load specific subset and split: from datasets import load_dataset # Load specific subset and split dataset = load_dataset('sujalappa/denoised-evaluation-dataset'… See the full description on the dataset page: https://huggingface.co/datasets/sujalappa/denoised-evaluation-dataset.audioautomatic-speech-recognitionn<1K0 likes41 downloads1y agoHugging Face14speech-uk /asr-evaluationstabularautomatic-speech-recognition10K<n<100K0 likes38 downloads2y agoHugging Face15MERA-evaluation /ruEnvAQA ruEnvAQA Описание задачи ruEnvAQA – датасет вопросов с множественным и бинарным выбором ответа на русском языке. Вопросы связаны с анализом музыки и невербальных аудиосигналов. Датасет составлен на основе вопросов из англоязычных датасетов Clotho-AQA и MUSIC-AVQA. Вопросы переведены на русский язык и частично изменены, тогда как аудиозаписи использованы в исходном виде (с обрезкой по длине). Датасет включает вопросы 8 типов: Оригинальные классы вопросов из MUSIC-AVQA… See the full description on the dataset page: https://huggingface.co/datasets/MERA-evaluation/ruEnvAQA.audion<1K0 likes37 downloads10mo agoHugging Face16sujalappa /speech-brain-noise-evaluation-dataset Speech Brain Noise Evaluation Dataset Dataset Description This dataset contains 2,000 samples organized across multiple splits and 20 subsets. The dataset includes audio data. Dataset Structure Subsets This dataset includes the following subsets: noisy-bg-snr-10: 100 samples test: 100 samples noisy-bg-snr-30: 100 samples test: 100 samples noisy-bg-snr-50: 100 samples test: 100 samples denoised-bg-snr-10: 100 samples test: 100 samples… See the full description on the dataset page: https://huggingface.co/datasets/sujalappa/speech-brain-noise-evaluation-dataset.audioautomatic-speech-recognition1K<n<10K0 likes35 downloads1y agoHugging Face17humanify /speaker_evaluation_multi_test_v0 Seamless Interaction Pairs This dataset contains paired query and document audio clips for interaction-based speaker evaluation. Each row describes a query clip and a related document clip, with segment metadata and durations for analysis. Data structure The dataset uses a single split stored in data.parquet. Audio files are stored under audio/ and referenced by relative paths in the parquet file. Columns pair_id (string): Pair identifier. interaction… See the full description on the dataset page: https://huggingface.co/datasets/humanify/speaker_evaluation_multi_test_v0.audioaudio-classificationn<1K0 likes35 downloads7mo agoHugging Face18isaacnetero /polyvox-kpi-evaluation PolyVox KPI Evaluation Dataset This dataset supports PolyVox evaluation for Problem Statement 11: Real-Time Multi-User Smart Assistant for Dynamic and Noisy Smart Environments. It contains compact evaluation mixtures for 2-speaker and 3-speaker speech separation under clean, noisy, overlap, and optional RIR-style conditions. Contents manifests/kpi_dataset_manifest.csv dataset_summary.json audio/ containing mixtures and clean source references Source… See the full description on the dataset page: https://huggingface.co/datasets/isaacnetero/polyvox-kpi-evaluation.audioautomatic-speech-recognition1K<n<10K0 likes34 downloads3mo agoHugging Face19MERA-evaluation /AQUARIA AQUARIA Описание задачи Датасет состоит из вопросов с выбором ответа, проверяющие комплексное понимание аудио, в том числе речи, неречевых сигналов и музыки. Вопросы датасета составлялись таким образом, чтобы для ответа на них требовалось не только распознавать речь, но и анализировать аудиоситуацию целиком и взаимодействие её компонентов. Используемые аудиофайлы созданы специально для датасета AQUARIA. В датасете представлены вопросы 9 типов: Audio scene classification… See the full description on the dataset page: https://huggingface.co/datasets/MERA-evaluation/AQUARIA.audion<1K0 likes28 downloads10mo agoHugging Face20humanify /speaker_evaluation_single_test_v0 Speaker Task Test Dataset 数据集描述 40 pairs sampled from Voxceleb1 118 pairs sampled from Voxceleb2 audioaudio-classificationn<1K0 likes21 downloads7mo agoHugging Face21MERA-evaluation /ruSLUn RuSLUn Описание задачи RuSLUn (Russian Spoken Language UNderstanding dataset) — это датасет для задачи понимания устной речи на русском языке, построенный по принципу англоязычного датасета SLURP и мультиязычного xSID, но с учетом культурных и языковых особенностей России. Он предназначен для оценки моделей, которые напрямую преобразуют аудиозаписи в семантическое представление, включая определение намерений пользователя (intent detection) и извлечение слотов (slot… See the full description on the dataset page: https://huggingface.co/datasets/MERA-evaluation/ruSLUn.audion<1K0 likes20 downloads10mo agoHugging Face22PhanithLIM /fleurs-evaluationaudion<1K0 likes16 downloads2y agoHugging Face23MERA-evaluation /ruTiE-Audio ruTiE-Audio Описание задачи ruTiE-Audio — мультимодальная эмуляция теста Тьюринга. Задача сформирована как неизменяемая последовательность вопросно-ответных заданий с опцией выбора ответа. Это 3 связных диалога, каждый с имитацией 500 обращений пользователя к модели. На вход модели подаётся аудио с заключёнными в аудиофайле заданиями и вопросами. Варианты ответа (4 к каждому заданию) модель получает текстом и выбирает из них. Задания теста проверяют способность модели… See the full description on the dataset page: https://huggingface.co/datasets/MERA-evaluation/ruTiE-Audio.audio1K<n<10K0 likes16 downloads10mo agoHugging Face24AhmedAshrafMarzouk /tts-evaluation-datasetaudion<1K0 likes16 downloads3mo agoHugging Face25keeve101 /fleurs-reducedbaseline-model-evaluationsaudion<1K0 likes15 downloads1y agoHugging Face26djopin /asr_evaluation_datasetsaudion<1K0 likes14 downloads11mo agoHugging Face27KeraCare /evaluation-whisper-large-v3-floresaudion<1K0 likes13 downloads2y agoHugging Face28prvInSpace /evaluation-setaudio1K<n<10K0 likes13 downloads1y agoHugging Face29wandererupak /nepali_asr_evaluation_dataaudio1K<n<10K0 likes12 downloads4mo agoHugging Face30iamTangsang /nepali_to_english_pipeline_evaluation Nepali-English Speech-to-Text Translation Evaluation Dataset Dataset Description This dataset is designed for evaluating Nepali→English speech-to-text translation pipelines. It contains audio recordings of 300 Nepali sentences, spoken by three speakers, covering a range of sentence types (statements, questions, commands, complex sentences, and named entities/numbers). Each sentence is paired with: Source text (Nepali) transcription Reference English translation Audio… See the full description on the dataset page: https://huggingface.co/datasets/iamTangsang/nepali_to_english_pipeline_evaluation.audiotranslationn<1K0 likes11 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.