datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
iamd_v0
Internet Archive Music Dataset (IAMD v0)
~4.2M thirty-second music segments (34,469 hours) sourced from
Creative-Commons audio on the Internet Archive, each
paired with machine-generated natural-language captions and the original item
metadata.
Segments
4.2M
Audio
34k hours
Segment length
30 s nominal (mean 29.22 s)
Format
MP3, 320 kbps CBR, native channels + sample rate
Shards
2,320 Parquet files
Download size
4.53 TB
Loading
A… See the full description on the dataset page: https://huggingface.co/datasets/Telecom-Paris/iamd_v0.bengali-telecom-customer-care-speech-v2
Bengali Telecom Customer Care Synthetic Speech Dataset v2
Dataset Description
This dataset contains synthetic Bengali speech generated from telecom and customer-care style text prompts.
The dataset is intended for experiments with:
Bengali ASR/STT
Bengali TTS
Speech-to-text preprocessing
Telecom/customer-care domain adaptation
Synthetic speech research
This is a second version of the Bengali Telecom Customer Care Synthetic Speech Dataset. It follows the same… See the full description on the dataset page: https://huggingface.co/datasets/kawshikbuet17/bengali-telecom-customer-care-speech-v2.bengali-telecom-customer-care-speech
Bengali Telecom Customer Care Synthetic Speech Dataset
Dataset Description
This dataset contains synthetic Bengali speech generated from telecom and customer-care style text prompts.
The dataset is intended for experiments with:
Bengali ASR/STT
Bengali TTS
Speech-to-text preprocessing
Telecom/customer-care domain adaptation
Synthetic speech research
Important Disclosure
This is a synthetic speech dataset generated using the OmniVoice TTS system in… See the full description on the dataset page: https://huggingface.co/datasets/kawshikbuet17/bengali-telecom-customer-care-speech.gsma-ethio-telecom-amharic-audiogsma-ethio-telecom-amharic-audio-v2gsma-ethio-telecom-amharic-audio-v2-mp3Telecom_Channel_Degredation_Matrix
SSA Codec Degradation Study — Acoustic Feature Exports
Moonscape Software | 2026
A companion to the Synthetic Speech Atlas (SSA)
Overview
This dataset quantifies the effect of 35 codec conditions on 80+ acoustic
features extracted from 7,500 biological speech clips. It answers the question:
"Which acoustic features survive telecommunications codec compression, and which
are destroyed?"
The corpus is the empirical foundation for channel-aware gate calibration in
deepfake… See the full description on the dataset page: https://huggingface.co/datasets/moonscape-software/Telecom_Channel_Degredation_Matrix.
