CoolFace
Datasetpublicgated

BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset

BabelSpeech 50-Hour Indonesian Colloquial ASR Speech Dataset Contains 50 hours of Indonesian colloquial ASR speech data, aligned with natural, everyday Indonesian communication patterns. Metadata is stored in a separate JSON file, including audio path, duration, transcript confidence, signal-to-noise ratio (SNR), and DNSMOS. More metadata fields may be added in future updates. Covered domains: technology, entertainment, travel, education, daily life, and others. Data quality:… See the full description on the dataset page: https://huggingface.co/datasets/BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset.

sourceHugging Faceapache-2.0updated 10mo agoView on Hugging Face
0likes18downloads

BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.