CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Telecom-Paris /iamd_v0 Internet Archive Music Dataset (IAMD v0) ~4.2M thirty-second music segments (34,469 hours) sourced from Creative-Commons audio on the Internet Archive, each paired with machine-generated natural-language captions and the original item metadata. Segments 4.2M Audio 34k hours Segment length 30 s nominal (mean 29.22 s) Format MP3, 320 kbps CBR, native channels + sample rate Shards 2,320 Parquet files Download size 4.53 TB Loading A… See the full description on the dataset page: https://huggingface.co/datasets/Telecom-Paris/iamd_v0.audioaudio-classification1M<n<10M6 likes2.1k downloads2mo agoHugging Face02kawshikbuet17 /bengali-telecom-customer-care-speech-v2 Bengali Telecom Customer Care Synthetic Speech Dataset v2 Dataset Description This dataset contains synthetic Bengali speech generated from telecom and customer-care style text prompts. The dataset is intended for experiments with: Bengali ASR/STT Bengali TTS Speech-to-text preprocessing Telecom/customer-care domain adaptation Synthetic speech research This is a second version of the Bengali Telecom Customer Care Synthetic Speech Dataset. It follows the same… See the full description on the dataset page: https://huggingface.co/datasets/kawshikbuet17/bengali-telecom-customer-care-speech-v2.audiotext-to-speech1K<n<10K0 likes70 downloads3mo agoHugging Face03kawshikbuet17 /bengali-telecom-customer-care-speech Bengali Telecom Customer Care Synthetic Speech Dataset Dataset Description This dataset contains synthetic Bengali speech generated from telecom and customer-care style text prompts. The dataset is intended for experiments with: Bengali ASR/STT Bengali TTS Speech-to-text preprocessing Telecom/customer-care domain adaptation Synthetic speech research Important Disclosure This is a synthetic speech dataset generated using the OmniVoice TTS system in… See the full description on the dataset page: https://huggingface.co/datasets/kawshikbuet17/bengali-telecom-customer-care-speech.audiotext-to-speech10K<n<100K0 likes22 downloads3mo agoHugging Face04gheero-Leyu /gsma-ethio-telecom-amharic-audiogatedaudio1K<n<10K0 likes7 downloads4mo agoHugging Face05gheero-Leyu /gsma-ethio-telecom-amharic-audio-v2gatedaudio10K<n<100K0 likes6 downloads3mo agoHugging Face06gheero-Leyu /gsma-ethio-telecom-amharic-audio-v2-mp3gatedaudio10K<n<100K1 likes6 downloads3mo agoHugging Face07moonscape-software /Telecom_Channel_Degredation_Matrixgated SSA Codec Degradation Study — Acoustic Feature Exports Moonscape Software | 2026 A companion to the Synthetic Speech Atlas (SSA) Overview This dataset quantifies the effect of 35 codec conditions on 80+ acoustic features extracted from 7,500 biological speech clips. It answers the question: "Which acoustic features survive telecommunications codec compression, and which are destroyed?" The corpus is the empirical foundation for channel-aware gate calibration in deepfake… See the full description on the dataset page: https://huggingface.co/datasets/moonscape-software/Telecom_Channel_Degredation_Matrix.tabularaudio-classification100K<n<1M0 likes5 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.