CoolFace
Datasetpublicgated

BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset

BabelSpeech 50-Hour Indonesian Colloquial ASR Speech Dataset Contains 50 hours of Indonesian colloquial ASR speech data, aligned with natural, everyday Indonesian communication patterns. Metadata is stored in a separate JSON file, including audio path, duration, transcript confidence, signal-to-noise ratio (SNR), and DNSMOS. More metadata fields may be added in future updates. Covered domains: technology, entertainment, travel, education, daily life, and others. Data quality:… See the full description on the dataset page: https://huggingface.co/datasets/BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
0likes18downloads

No commit history came back for main. The revision may not exist, or the source declined the request.