CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01WhissleAI /Meta_STT_HI_Set1 Meta Speech Recognition Hindi Dataset (Set 1) This dataset contains both metadata and audio files for Hindi speech recognition samples, curated from multiple sources. Dataset Sources and Credits This dataset combines samples from the following sources: AI4Bharat Indic Speech Dataset Source: https://ai4bharat.org/indic-speech-dataset License: CC-BY 4.0 Citation: Please cite the original paper if you use this data Common Voice Hindi Source:… See the full description on the dataset page: https://huggingface.co/datasets/WhissleAI/Meta_STT_HI_Set1.audioautomatic-speech-recognition100K<n<1M0 likes467 downloads1y agoHugging Face02WhissleAI /Meta_STT_ZH_AIShell3 Meta Speech Recognition Mandarin Dataset (AISHELL3) This dataset contains both metadata and audio files for Mandarin speech recognition samples from the AISHELL3 corpus. Dataset Statistics Splits and Sample Counts train: 60098 samples valid: 3163 samples test: 24772 samples Example Samples train { "audio_filepath": "/external4/datasets/Mandarin/AISHELL3/wavs_train/SSB00430356.wav", "text": "她以 ENTITY_PRODUCT 滴鸡精 END 调养身体。 AGE_14_25… See the full description on the dataset page: https://huggingface.co/datasets/WhissleAI/Meta_STT_ZH_AIShell3.audioautomatic-speech-recognition10K<n<100K0 likes465 downloads1y agoHugging Face03WhissleAI /Meta_STT_EN_Set2 Meta Speech Recognition English Dataset (Set 2) This dataset contains both metadata and audio files for English speech recognition samples. Dataset Statistics Splits and Sample Counts train: 42961 samples valid: 2387 samples test: 2387 samples Example Samples train { "audio_filepath": "/external1/datasets/asr-himanshu/avspeech-data/audio/AzSutepklXI_2.wav", "text": "To Jesus, so God is faithful, because when he keeps, you know, when… See the full description on the dataset page: https://huggingface.co/datasets/WhissleAI/Meta_STT_EN_Set2.audioaudio-classification10K<n<100K0 likes340 downloads1y agoHugging Face04WhissleAI /betrac-2026-with-meta betrac-2026-with-meta Annotated speech dataset created with Whissle Annotator — a multimodal annotation pipeline for speech, NLP, and visual analysis. Source Dataset This dataset is derived from the following HuggingFace dataset(s): BeTraC/betrac-2026:0:50 BeTraC/betrac-2026:50:50 BeTraC/betrac-2026:100:50 BeTraC/betrac-2026:150:50 BeTraC/betrac-2026:200:50 BeTraC/betrac-2026:250:50 BeTraC/betrac-2026:300:50 BeTraC/betrac-2026:350:50 BeTraC/betrac-2026:400:50… See the full description on the dataset page: https://huggingface.co/datasets/WhissleAI/betrac-2026-with-meta.audioautomatic-speech-recognition1K<n<10K0 likes39 downloads5mo agoHugging Face05WhissleAI /Meta_STT_MADASR2.0_train_lggated Dataset Card for Dataset Name https://sites.google.com/view/respinasrchallenge2025/home?authuser=0 This is the trainig dataset provided for track-3 and track-4 of this task. We enhance the dataset with entity tagging, emotion, age, gender and intent. WhissleAI participated in this challenge. This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated… See the full description on the dataset page: https://huggingface.co/datasets/WhissleAI/Meta_STT_MADASR2.0_train_lg.audioaudio-classification100K<n<1M0 likes9 downloads4mo agoHugging Face06WhissleAI /Meta_STT_EN-IN_Tech_Interviewsgated Meta STT EN-IN Tech Interviews Indian English speech recognition dataset sourced from technical interviews, annotated with rich speech metadata including age group, gender, emotion, and intent. Designed for training multi-task ASR models that jointly predict transcriptions and speaker attributes. Dataset Details Property Value Train examples 58,000 Validation examples 1,204 Language English (Indian accent) Audio 16 kHz Total size ~28 GB… See the full description on the dataset page: https://huggingface.co/datasets/WhissleAI/Meta_STT_EN-IN_Tech_Interviews.audioautomatic-speech-recognition10K<n<100K0 likes8 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.