CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01devrahulbanjara /ne-en-codeswitching-asr-technical-interview Dataset Summary This dataset contains audio recordings and text transcripts of Nepali-English code-switched speech in the context of technical interviews. It is specifically designed to handle the linguistic complexities of Nepali software engineers, developers, and IT professionals who frequently mix English technical terminology (e.g., AWS, S3 lifecycle policies, RAG pipelines, VPC peering) with conversational Nepali grammar. It is an excellent resource for fine-tuning ASR models… See the full description on the dataset page: https://huggingface.co/datasets/devrahulbanjara/ne-en-codeswitching-asr-technical-interview.audioautomatic-speech-recognitionn<1K3 likes102 downloads7mo agoHugging Face02Yassmen /TTS_English_Technical_dataaudio1K<n<10K3 likes26 downloads2y agoHugging Face03Shabdobhedi /TTS_English_Technical_Termsaudio1K<n<10K0 likes26 downloads2y agoHugging Face04WpythonW /elevenlabs_multilingual_v2-technical-speech ElevenLabs Multilingual V2 Technical Speech Dataset This dataset contains automatically generated technical phrases in three domains, converted to speech using the ElevenLabs Multilingual V2 model with Adam voice. Dataset Description The dataset includes audio samples of technical phrases across three categories: Machine Learning (ML) Science Technology Each entry contains: Audio file in MP3 format (22050Hz) Source text Text length Category label Data… See the full description on the dataset page: https://huggingface.co/datasets/WpythonW/elevenlabs_multilingual_v2-technical-speech.audiotext-to-speechn<1K0 likes14 downloads2y agoHugging Face05Tejasva-Maurya /English-Technical-Speech-Datasetgated English Technical Speech Dataset Overview The English Technical Speech Dataset is a curated collection of English technical vocabulary recordings, designed for applications like Text-to-Speech (TTS), Automatic Speech Recognition (ASR), and Audio Classification. The dataset includes 11,247 entries and provides audio files, transcriptions, and speaker embeddings to support the development of robust technical language models. Language: English (technical focus) Total… See the full description on the dataset page: https://huggingface.co/datasets/Tejasva-Maurya/English-Technical-Speech-Dataset.audio10K<n<100K8 likes7 downloads2y agoHugging Face06Vinay15 /Technical_Terms_With_Pronunciations_and_Audiogatedaudion<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.