CoolFace
3 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Appenlimited /1000h-us-english-smartphone-conversation 📚 1000 Hours of Conversational American English Speech Dataset (Smartphone Recordings) This dataset contains sample conversational speech data collected by Appen. The audio was recorded naturally using smartphones and is suitable for: Automatic Speech Recognition (ASR) Speaker Identification and Gender/Age Analysis Dialect and Accent Modeling Multi-speaker Speech Separation 🧾 Dataset Contents The dataset includes: metadata.CSV: Metadata including speaker gender, age… See the full description on the dataset page: https://huggingface.co/datasets/Appenlimited/1000h-us-english-smartphone-conversation.audioautomatic-speech-recognitionn<1K3 likes138 downloads1y agoHugging Face02Porameht /processed-smarthome-th processed-smarthome-th Cleaned Thai speech dataset for smart-home commands: 9,600 utterances (7,680 train / 960 dev / 960 test) with transcripts. Format Field Description sentence Transcript in Thai audio Audio clip Usage from datasets import load_dataset ds = load_dataset("Porameht/processed-smarthome-th") Used to fine-tune Porameht/whisper-tiny-smarthome-thai (WER 24.375 on the eval split). audioautomatic-speech-recognition1K<n<10K1 likes33 downloads23d agoHugging Face03jurgenpaul82 /smartsteinfrom datasets import load_dataset ds = load_dataset("nyu-mll/glue", "ax") from datasets import load_dataset ds = load_dataset("nyu-mll/glue", "cola") from datasets import load_dataset ds = load_dataset("nyu-mll/glue", "mnli") text-classification100M<n<1B0 likes17 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.